← Model list

Llama 3.1 Nemotron 70B Instruct

nvidia·nvidia/llama-3.1-nemotron-70b-instruct·GA·Open weights·nemotron series

Nemotron model for efficient reasoning, coding, and specialized AI agents

At a glance
Official price$0 / $0 per 1M
Cache read · Nvidia
Lowest paid$0.60 / $0.60
Eden AI Gateway
1 more $0 channels
Context128,000
⚠ Providers report 128,000–131,072; the table below is authoritative
Output limit8,192
Capabilities
ReasoningTool useStructured outputTemperatureAttachments
Modalities
Text
Knowledge cutoff
Released / updated2025-04-15 / 2025-04-15

Quality & performance

Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.

Available at 2 providers1 with public prices · 1 free

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NvidiaOfficialFirst-partyFree128,0008,192
Eden AI
deepinfra/nvidia/Llama-3.1-Nemotron-70B-Instruct
Gateway$0.60$0.60131,0728,192

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Eden AI$150.00

1 more channels offer $0 (Nvidia); free tiers usually have rate limits and no SLA, excluded from ranking.

The cheapest paid channel is the only channel.

Related models

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.60$02026-08-052026-08-152026-08-05 · Input list $02026-08-06 · Input list $02026-08-07 · Input list $02026-08-08 · Input list $02026-08-09 · Input list $02026-08-10 · Input list $02026-08-11 · Input list $02026-08-12 · Input list $02026-08-13 · Input list $02026-08-15 · Input list $02026-08-05 · Output list $02026-08-06 · Output list $02026-08-07 · Output list $02026-08-08 · Output list $02026-08-09 · Output list $02026-08-10 · Output list $02026-08-11 · Output list $02026-08-12 · Output list $02026-08-13 · Output list $02026-08-15 · Output list $02026-08-15 · Min blended $0.60