← Model list
NVIDIA Nemotron Cascade 2
nvidia·nvidia/nemotron-cascade-2-30b-a3b·GA·Open weights·nemotron series
Nemotron model for efficient reasoning, coding, and specialized AI agents
At a glance
Reference price$0.15 / $0.60 per 1M
Cache read — · Vultr
Lowest paid$0.15 / $0.60
Vultr Cloud
Context262,144
Output limit131,072
Capabilities
✓ Reasoning✓ Tool use? Structured output✓ TemperatureAttachments
Modalities
Text
Knowledge cutoff2024-07
Released / updated2025-12-01 / 2025-12-01
Quality & performanceArtificial Analysis · Intelligence Index v4.1
Intelligence17.8
Value68
Coding25.3
Agentic—
Output speed—
TTFT—
Value formulaIQ 17.8 ÷ min blended $0.263 = 68
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 1 providers1 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Vultr nvidia/Nemotron-Cascade-2-30B-A3B | Cloud | $0.15 | $0.60 | — | — | 262,144 | 131,072 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Vultr$60.00
The cheapest paid channel is the only channel.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
Upstream provides no control info
1 / 1 providers expose no reasoning control (reasoning_options: []).
Related models
nemotron-lightning-3.5-30b-a3bsame series$0.045 / $0.18Nemotron 3.5 Lightning 30B A3Bsame series$0 / $0Nvidia Nemotron 3.5 Lightningsame series$0.05 / $0.20GPT-5 Nanocheaper alternative$0.05 / $0.40GLM-4.7-Flashcheaper alternative$0 / $0Step 3.5 Flashcheaper alternative$0.10 / $0.30Hy3cheaper alternative$0 / $0
Price historyone sample accumulated per data sync
Input listOutput listMin blended