← Model list
nvidia-nemotron-3-nano-omni
nvidia·nvidia/nvidia-nemotron-3-nano-omni·GA·Open weights·nemotron series
Open Nemotron omni model combining reasoning with text, vision, and audio
At a glance
Reference price$0.13 / $0.38 per 1M
Cache read — · Vultr
Lowest paid$0.059 / $0.237
Cortecs Gateway · 2.2× spread
Context262,144
⚠ Providers report 256,000–300,000; the table below is authoritative
Output limit131,072
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
TextImage
Knowledge cutoff2025-05
Released / updated2026-04-28 / 2026-04-29
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 3 providers3 with public prices
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Cortecs$23.65
2NanoGPT$35.70
3Vultr$45.00
The cheapest paid channel is the only channel.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
Upstream provides no control info
3 / 3 providers expose no reasoning control (reasoning_options: []).
Related models
nemotron-lightning-3.5-30b-a3bsame series$0.045 / $0.18Nemotron 3.5 Lightning 30B A3Bsame series$0 / $0Nvidia Nemotron 3.5 Lightningsame series$0.05 / $0.20GLM-4.7-Flashcheaper alternative$0 / $0Hy3cheaper alternative$0 / $0Nemotron 3 Nano 30B A3Bcheaper alternative$0 / $0Qwen3.7 Flashcheaper alternative$0.03 / $0.118
Price historyone sample accumulated per data sync
Input list $0.13Output list $0.38Min blended $0.103
Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-13). Each future sync adds a point, and once accumulated a line is drawn here.