← Model list
Nemotron 3 Nano Omni
nvidia·nvidia/nemotron-3-nano-omni·GA·Open weights·nemotron series
Open Nemotron omni model combining reasoning with text, vision, and audio
At a glance
Official price$0 / $0 per 1M
Cache read — · Nvidia
Lowest paid$0.054 / $0.216
Requesty Gateway · 9.3× spread
1 more $0 channels
Context256,000
⚠ Providers report 65,536–300,000; the table below is authoritative
Output limit65,536
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImageVideoAudio
Knowledge cutoff2025-01
Released / updated2026-04-28 / 2026-05-20
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 5 providers4 with public prices · 1 free
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NvidiaOfficial nvidia/nemotron-3-nano-omni-30b-a3b-reasoning | First-party | Free | — | — | 256,000 | 65,536 | ||
| Requesty | Gateway | $0.054 | $0.216 | $0.054 | — | 300,000 ⚠ | 300,000 | |
| Nebius Token Factory nvidia/Nemotron-3-Nano-Omni | Cloud | $0.06 | $0.24 | $0.006 | $0.075 | 65,536 ⚠ | 8,192 | |
| DevPass (LLM Gateway) | Gateway | $0.06 | $0.24 | — | — | 262,144 ⚠ | 262,144 | |
| DigitalOcean | Cloud | $0.50 | $0.90 | — | — | 65,536 ⚠ | 65,536 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Nebius Token Factory$17.52
2Requesty$21.60
3DevPass (LLM Gateway)$24.00
4DigitalOcean$145.00
1 more channels offer $0 (Nvidia); free tiers usually have rate limits and no SLA, excluded from ranking.
The cheapest paid channel is the only channel.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
Toggle (on / off)budget_tokenseffort = noneeffort = loweffort = mediumeffort = higheffort = max
1 / 5 providers expose no reasoning control (reasoning_options: []).
Related models
nemotron-lightning-3.5-30b-a3bsame series$0.045 / $0.18Nemotron 3.5 Lightning 30B A3Bsame series$0 / $0Nvidia Nemotron 3.5 Lightningsame series$0.05 / $0.20
Price historyone sample accumulated per data sync
Input listOutput listMin blended