← Model list
Llama 3.3 Nemotron Super 49B v1
nvidia·nvidia/llama-3.3-nemotron-super-49b-v1·GA·Open weights·nemotron series
Nemotron model for efficient reasoning, coding, and specialized AI agents
At a glance
Official price$0 / $0 per 1M
Cache read — · Nvidia
Lowest paidNo paid channels
1 more $0 channels
Context131,072
Output limit65,536
Capabilities
✓ Reasoning✓ Tool use? Structured output✓ TemperatureAttachments
Modalities
Text
Knowledge cutoff—
Released / updated2025-04-07 / 2025-04-07
Quality & performanceArtificial Analysis · Intelligence Index v4.1 · rep. tier Reasoning
Intelligence12.2
Value—
Coding—
Agentic—
Output speed—
TTFT—
Reasoning tier → intelligence / speed (higher tier = stronger but slower)
Non-reasoningIQ 8.3
ReasoningIQ 12.2
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 1 providers0 with public prices · 1 free
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NvidiaOfficial | First-party | Free | — | — | 131,072 | 65,536 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
This model is $0 across all listed channels (Nvidia).
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
Toggle (on / off)
Related models
nemotron-lightning-3.5-30b-a3bsame series$0.045 / $0.18Nemotron 3.5 Lightning 30B A3Bsame series$0 / $0Nvidia Nemotron 3.5 Lightningsame series$0.05 / $0.20
Price historyone sample accumulated per data sync
Input listOutput listMin blended