← Model list

Nvidia Nemotron 3.5 Lightning

nvidia·nvidia/nemotron-3.5-lightning-2·GA·Open weights·nemotron series·NEW

Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads

At a glance
Reference price$0.05 / $0.20 per 1M
Cache read $0.01 · NanoGPT
Lowest paid$0.05 / $0.20
NanoGPT Gateway
Context1,000,000
Output limit65,536
Capabilities
ReasoningTool useStructured outputTemperatureAttachments
Modalities
Text
Knowledge cutoff
Released / updated2026-08-11 / 2026-08-11

Quality & performance

Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NanoGPT
nvidia/nemotron-3.5-lightning
Gateway$0.05$0.20$0.011,000,00065,536

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1NanoGPT$15.20
The cheapest paid channel is the only channel.

Benchmark

No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.

Reasoning control

effort = noneeffort = high

Related models

Price historyone sample accumulated per data sync

Input list $0.05Output list $0.20Min blended $0.087

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-13). Each future sync adds a point, and once accumulated a line is drawn here.