← Model list

nvidia-nemotron-3-nano-omni

nvidia·nvidia/nvidia-nemotron-3-nano-omni·GA·Open weights·nemotron series

Open Nemotron omni model combining reasoning with text, vision, and audio

At a glance
Reference price$0.13 / $0.38 per 1M
Cache read · Vultr
Lowest paid$0.059 / $0.237
Cortecs Gateway · 2.2× spread
Context262,144
⚠ Providers report 256,000–300,000; the table below is authoritative
Output limit131,072
Capabilities
ReasoningTool useStructured outputTemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
TextImage
Knowledge cutoff2025-05
Released / updated2026-04-28 / 2026-04-29

Quality & performance

Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.

Available at 3 providers3 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
CortecsGateway$0.059$0.237300,000300,000
NanoGPT
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning
Gateway$0.105$0.42$0.052256,00065,536
Vultr
nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16
Cloud$0.13$0.38262,144131,072

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Cortecs$23.65
2NanoGPT$35.70
3Vultr$45.00
The cheapest paid channel is the only channel.

Benchmark

No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.

Reasoning control

Upstream provides no control info

3 / 3 providers expose no reasoning control (reasoning_options: []).

Related models

Price historyone sample accumulated per data sync

Input list $0.13Output list $0.38Min blended $0.103

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-13). Each future sync adds a point, and once accumulated a line is drawn here.