← Model list
Mistral Small 3.1
mistral·mistral/mistral-small-2503·GA·Closed·mistral-small series
Efficient Mistral model for fast chat, extraction, and production assistants
At a glance
Reference price$0.10 / $0.30 per 1M
Cache read — · Azure
Lowest paid$0.10 / $0.30
Azure Cognitive Services First-party · 1× spread
Context128,000
Output limit32,768
Capabilities
Reasoning✓ Tool use? Structured output✓ Temperature✓ Attachments
Modalities
TextImage
Knowledge cutoff2024-09
Released / updated2025-03-01 / 2025-03-01
Quality & performanceArtificial Analysis · Intelligence Index v4.1
Intelligence14.9
Value99
Coding26.3
Agentic5.3
Output speed149 tok/s
TTFT0.81 s
Cost per task$0.0443
Value formulaIQ 14.9 ÷ min blended $0.150 = 99
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 2 providers2 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Azure Cognitive Services | First-party | $0.10 | $0.30 | — | — | 128,000 | 32,768 | |
| Azure | First-party | $0.10 | $0.30 | — | — | 128,000 | 32,768 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Azure Cognitive Services$35.00
2Azure$35.00
The cheapest paid channel is the only channel.
Related models
Mistral Small 3.2 24B Instructsame series$0.20 / $0.40mistral-small-3.2-24b-instruct-2506same series$0.33 / $0.33Mistral Small 4 119B Thinkingsame series$0.40 / $1.40Llama 4 Maverick 17B 128E Instruct FP8cheaper alternative$0 / $0Lyria 3 Clip Previewcheaper alternative$0 / $0Command R7Bcheaper alternative$0.037 / $0.15Granite 4.1 8Bcheaper alternative$0.05 / $0.10
Price historyone sample accumulated per data sync
Input listOutput listMin blended