LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Nvidia Nemotron 3.5 Lightning Thinking

nvidia·nvidia/nemotron-3.5-lightning-thinking·GA·Open weights·nemotron series·NEW
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads

Specs & pricing

Input / output per 1M tokens
Reference price·NanoGPT
$0.05 / $0.20
Blended $0.088 · Cache read $0.01
Lowest paid·NanoGPTGateway
$0.05 / $0.20
Blended $0.088
Context
1,000,000
Output limit
65,536
Knowledge cutoff
—
Released / updated
2026-08-11 / 2026-08-11
Capabilities
✓ Reasoning✓ Tool useStructured output✓ TemperatureAttachments
Modalities
Text

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NanoGPT
nvidia/nemotron-3.5-lightning:thinking
Gateway$0.05$0.20$0.01—1,000,00065,536

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = noneeffort = high

Your usage cost

1NanoGPT$15.20
The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input list $0.05Output list $0.20Min blended $0.088

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-13). Each future sync adds a point, and once accumulated a line is drawn here.

Related models

Nemotron 3.5 Lightning 30B A3Bsame series$0 / $0Nemotron 3.5 Lightning (free)same series$0 / $0Nemotron 3.5 Lightningsame series$0.07 / $0.20Laguna S 2.1cheaper alternative$0 / $0Nemotron 3 Ultra (free)cheaper alternative$0 / $0miAuto Routercheaper alternative$0 / $0
Data partly from models.dev (MIT) · NanoGPT official docs ↗