Skip to content

Nemotron 3.5 Lightning 30B A3BFree

Also accepted:nvidia/nemotron-3.5-lightningnvidia/nemotron-3.5-lightning:free

NVIDIA Nemotron 3.5 Lightning 30B A3B is a sparse MoE chat model optimized for low-latency agentic workloads on NVIDIA NIM.

Providers
Capabilities
NVIDIA
nvidia
$0
$0
AIHubMix
aihubmix-byok
$0
$0
Unavailable
NVIDIA
nvidia-byok
$0
$0
Unavailable
OpenRouter
openrouter-byok
$0
$0
Unavailable
Usage analytics

Loading usage…

API & code
Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Share cards
Nemotron 3.5 Lightning 30B A3B share card
Nemotron 3.5 Lightning 30B A3B
AIHubMix upstream share card
AIHubMix upstream
NVIDIA upstream share card
NVIDIA upstream
NVIDIA upstream share card
NVIDIA upstream
Hue upstream share card
Hue upstream
OpenRouter upstream share card
OpenRouter upstream
Credits
Use your own key

Run Nemotron 3.5 Lightning 30B A3B on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Text generation
Context length1,000,000 tokens
Max output65,536 tokens
ArchitectureTransformer
Categorytext
ReleasedAug 1, 2026
Modalities
Capabilities