← Model list

Llama 3.3 Nemotron Super 49B v1

nvidia·nvidia/llama-3.3-nemotron-super-49b-v1·GA·Open weights·nemotron series

Nemotron model for efficient reasoning, coding, and specialized AI agents

At a glance
Official price$0 / $0 per 1M
Cache read · Nvidia
Lowest paidNo paid channels
1 more $0 channels
Context131,072
Output limit65,536
Capabilities
ReasoningTool use? Structured outputTemperatureAttachments
Modalities
Text
Knowledge cutoff
Released / updated2025-04-07 / 2025-04-07

Quality & performanceArtificial Analysis · Intelligence Index v4.1 · rep. tier Reasoning

Intelligence12.2
Value
Coding
Agentic
Output speed
TTFT
Reasoning tier → intelligence / speed (higher tier = stronger but slower)
Non-reasoningIQ 8.3
ReasoningIQ 12.2

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 1 providers0 with public prices · 1 free

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NvidiaOfficialFirst-partyFree131,07265,536

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

This model is $0 across all listed channels (Nvidia).

Benchmark

No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.

Reasoning control

Toggle (on / off)

Related models

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.0001$02026-08-052026-08-132026-08-05 · Input list $02026-08-06 · Input list $02026-08-07 · Input list $02026-08-08 · Input list $02026-08-09 · Input list $02026-08-10 · Input list $02026-08-11 · Input list $02026-08-12 · Input list $02026-08-13 · Input list $02026-08-05 · Output list $02026-08-06 · Output list $02026-08-07 · Output list $02026-08-08 · Output list $02026-08-09 · Output list $02026-08-10 · Output list $02026-08-11 · Output list $02026-08-12 · Output list $02026-08-13 · Output list $0