← Model list
Gemini 3.1 Pro Preview
google·google/gemini-3.1-pro-preview·GA·Closed·gemini-pro series
Reasoning-first Gemini preview for agentic coding and complex problem solving
At a glance
Official price$2.00 / $12.00 per 1M
Cache read $0.20 · Vertex
Lowest paid$1.80 / $10.80
Requesty Gateway · 2.2× spread
Context1,048,576
⚠ Providers report 1,000,000–1,048,756; the table below is authoritative
Output limit65,536
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImageVideoAudioPDF
Knowledge cutoff2025-01
Released / updated2026-02-19 / 2026-02-19
Quality & performanceArtificial Analysis · Intelligence Index v4.1
Intelligence47.7
Value12
Coding68.8
Agentic23
Output speed108 tok/s
TTFT32.45 s
Cost per task$0.3346
Value formulaIQ 47.7 ÷ min blended $4.050 = 12
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 31 providers30 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Requesty | Gateway | $1.80 tiered >200K: $3.60 | $10.80 | $0.18 | $4.05 | 1,048,576 | 65,535 | |
| NanoGPT | Gateway | $2.00 | $12.00 | $0.20 | $0.375 | 1,048,756 ⚠ | 65,536 | |
| VertexOfficial | First-party | $2.00 tiered >200K: $4.00 | $12.00 | $0.20 | — | 1,048,576 | 65,536 | |
| GoogleOfficial | First-party | $2.00 tiered >200K: $4.00 | $12.00 | $0.20 | — | 1,048,576 | 65,536 | |
| Impossibl | Gateway | $2.00 tiered >200K: $4.00 | $12.00 | $0.20 | — | 1,048,576 | 65,536 | |
| OpenRouter | Gateway | $2.00 tiered >200K: $4.00 | $12.00 | $0.20 | $0.375 | 1,048,576 | 65,536 | |
| Auriko | Gateway | $2.00 tiered >200K: $4.00 | $12.00 | $0.20 | — | 1,048,576 | 65,536 | |
| DaoXE | Gateway | $2.00 | $12.00 | $0.20 | — | 1,048,576 | 65,536 | |
| CrossModel | Gateway | $2.00 tiered >200K: $4.00 | $12.00 | $0.20 | $2.00 | 1,048,576 | 65,536 | |
| Vivgrid | Gateway | $2.00 tiered >200K: $4.00 | $12.00 | $0.20 | — | 1,048,576 | 65,536 | |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Requesty$705.60
2NanoGPT$784.00
3Vertex · Official$784.00
4Google · Official$784.00
5Impossibl$784.00
6OpenRouter$784.00
Switch to Requesty to save $78.40/mo (10%). The gap is small, so staying on the official channel is fine.
Note: this is a gateway; verify availability and rate limits yourself.
Benchmark18 items
| Name | Conditions | Score | Metric | Source |
|---|---|---|---|---|
| SWE-Bench Pro | — | 54.2 | resolve rate | Source ↗ |
| Terminal-Bench | harness: Terminus-2 · v2.1 | 70.3 | success rate | Source ↗ |
| SWE-Bench Pro | dataset: public | 46.1 | resolve rate | Source ↗ |
| SWE-Atlas Codebase QnA | harness: Mini-SWE-Agent | 13.5 | score | Source ↗ |
| SWE-Atlas Refactoring | harness: Gemini CLI | 33.81 | score | Source ↗ |
| SWE-Atlas Test Writing | harness: Mini-SWE-Agent | 29.84 | score | Source ↗ |
| Artificial Analysis Coding Agent Index | harness: Gemini CLI · variant: high | 43 | average pass@1 | Source ↗ |
| SWE-Atlas Codebase QnA | harness: Gemini CLI · variant: high | 45.6 | pass@1 | Source ↗ |
| SWE-Bench Pro | harness: Gemini CLI · variant: high · dataset: hard-aa | 15.1 | pass@1 | Source ↗ |
| Terminal-Bench | harness: Gemini CLI · variant: high · v2.1 | 68.3 | pass@1 | Source ↗ |
| GPQA Diamond | — | 94.3 | accuracy | Source ↗ |
| Humanity's Last Exam | dataset: full set, text + MM | 44.4 | accuracy | Source ↗ |
| ARC-AGI-2 | — | 77.1 | accuracy | Source ↗ |
| MMMU Pro | variant: no tools | 80.5 | accuracy | Source ↗ |
| MCP Atlas | — | 78.2 | success rate | Source ↗ |
| OSWorld-Verified | — | 76.2 | success rate | Source ↗ |
| CharXiv Reasoning | variant: no tools | 83.3 | accuracy | Source ↗ |
| GDPval-AA | — | 1314 | Elo | Source ↗ |
The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.
Reasoning control
effort = noneeffort = loweffort = mediumeffort = higheffort = maxbudget_tokens ≥ 256 ≤ 32,000effort = xhigh
2 / 31 providers expose no reasoning control (reasoning_options: []).
Related models
Gemini 3.5 Live Translate Previewsame series$3.50 / $21.00Gemini 3 Pro Imagesame series$2.00 / $12.00Nano Banana Prosame series$2.00 / $120GLM-5.2cheaper alternative$1.40 / $4.40DeepSeek V4 Procheaper alternative$0.435 / $0.87DeepSeek V4 Flashcheaper alternative$0.14 / $0.28DeepSeek V4 Flash 0731cheaper alternative$0.479 / $1.44
Price historyone sample accumulated per data sync
Input listOutput listMin blended