LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

nvidia-nemotron-3-ultra

nvidia·nvidia/nvidia-nemotron-3-ultra·GA·Closed·nemotron series
NVIDIA Nemotron 3 Ultra is NVIDIA's strongest open-weights reasoning model, positioned near GPT-5.4 Mini (xhigh) and ahead of DeepSeek V4-Flash and Qwen3.5-397B-A17B.

nvidia-nemotron-3-ultra is currently listed from a single provider. Its reference price is $0.50 per 1M input tokens and $2.50 per 1M output tokens.

The context window is 262,144 tokens, with an output limit of 131,072 tokens. It supports reasoning.

Specs & pricing

Input / output per 1M tokens
Reference price·Requesty
$0.50 / $2.50
Blended $1.00 · Cache read —
Lowest paid·RequestyGateway
$0.50 / $2.50
Blended $1.00
Context
262,144
Output limit
131,072
Knowledge cutoff
—
Released / updated
2026-06-23 / 2026-06-23
Capabilities
✓ ReasoningTool useStructured output? TemperatureAttachments
Modalities
Text

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
RequestyGateway$0.50$2.50——262,144131,072

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = noneeffort = loweffort = mediumeffort = higheffort = max

Your usage cost

1Requesty$225.00
The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$3.13$02026-08-192026-10-112026-08-19 · Input list $0.632026-08-21 · Input list $0.632026-08-28 · Input list $0.632026-10-11 · Input list $0.502026-08-19 · Output list $3.132026-08-21 · Output list $3.132026-08-28 · Output list $3.132026-10-11 · Output list $2.502026-08-19 · Min blended $1.002026-08-21 · Min blended $0.902026-08-28 · Min blended $1.002026-10-11 · Min blended $1.00

Related models

Nemotron 3.5 Lightning 30B A3Bsame series$0 / $0Nemotron 3 Nano Omni 30B TEEsame series$0.025 / $0.098nemotron-3-ultra-nvfp4same series$0.60 / $2.40GLM-5.3-Flashcheaper alternative$0.15 / $0.50GPT-5.6 Lunacheaper alternative$0.20 / $1.20Qwen3.6 35B-A3Bcheaper alternative$0.25 / $1.49GPT-5.4 nanocheaper alternative$0.20 / $1.25
Data partly from models.dev (MIT) · Requesty official docs ↗