← Model list

sarvam-m

sarvamai·sarvamai/sarvam-m·GA·Open weights

Efficient Indian-language reasoning model for chat, coding, and multilingual work

At a glance
Reference price$0 / $0 per 1M
Cache read · Nvidia
Lowest paidNo paid channels
1 more $0 channels
Context128,000
Output limit8,192
Capabilities
ReasoningTool use? Structured outputTemperatureAttachments
Modalities
Text
Knowledge cutoff
Released / updated2025-07-25 / 2025-07-25

Quality & performanceArtificial Analysis · Intelligence Index v4.1 · rep. tier Reasoning

Intelligence2.6
Value
Coding
Agentic
Output speed
TTFT

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 1 providers0 with public prices · 1 free

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NvidiaFirst-partyFree128,0008,192

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

This model is $0 across all listed channels (Nvidia).

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.0001$02026-08-052026-08-132026-08-05 · Input list $02026-08-06 · Input list $02026-08-07 · Input list $02026-08-08 · Input list $02026-08-09 · Input list $02026-08-10 · Input list $02026-08-11 · Input list $02026-08-12 · Input list $02026-08-13 · Input list $02026-08-05 · Output list $02026-08-06 · Output list $02026-08-07 · Output list $02026-08-08 · Output list $02026-08-09 · Output list $02026-08-10 · Output list $02026-08-11 · Output list $02026-08-12 · Output list $02026-08-13 · Output list $0