← Model list
Gemini 3 Flash Preview
google·google/gemini-3-flash-preview·GA·Closed·gemini-flash series
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
At a glance
Official price$0.50 / $3.00 per 1M
Cache read $0.05 · Vertex
Lowest paid$0.07 / $0.43
QiHang Gateway · 10× spread
Context1,048,576
⚠ Providers report 256,000–1,048,756; the table below is authoritative
Output limit65,536
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImageVideoAudioPDF
Knowledge cutoff2025-01
Released / updated2025-12-17 / 2025-12-17
Quality & performanceArtificial Analysis · Intelligence Index v4.1 · rep. tier Reasoning
Intelligence38.7
Value242
Coding—
Agentic—
Output speed161 tok/s
TTFT6.22 s
Value formulaIQ 38.7 ÷ min blended $0.160 = 242
Reasoning tier → intelligence / speed (higher tier = stronger but slower)
Non-reasoningIQ 27.9 · 171 tok/s
ReasoningIQ 38.7 · 161 tok/s
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 25 providers24 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| QiHang | Gateway | $0.07 tiered >200K: $0.07 | $0.43 | — | — | 1,048,576 | 65,536 | |
| OrcaRouter | Gateway | $0.50 | $3.00 | $0.05 | — | 1,048,576 | 65,536 | |
| NanoGPT | Gateway | $0.50 | $3.00 | $0.05 | — | 1,048,756 ⚠ | 65,536 | |
| VertexOfficial | First-party | $0.50 | $3.00 | $0.05 | — | 1,048,576 | 65,536 | |
| GoogleOfficial | First-party | $0.50 | $3.00 | $0.05 | — | 1,048,576 | 65,536 | |
| Jiekou.AI | Gateway | $0.50 | $3.00 | — | — | 1,048,576 | 65,536 | |
| OpenRouter | Gateway | $0.50 | $3.00 | $0.05 | $0.083 | 1,048,576 | 65,536 | |
| CrossModel | Gateway | $0.50 | $3.00 | $0.05 | $0.50 | 1,048,576 | 65,536 | |
| Databricks databricks-gemini-3-flash | Cloud | $0.50 | $3.00 | $0.05 | — | 1,048,576 | 65,536 | |
| Merge Gateway | Gateway | $0.50 | $3.00 | $0.05 | — | 1,048,576 | 65,536 | |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1QiHang$35.50
2OrcaRouter$196.00
3NanoGPT$196.00
4Vertex · Official$196.00
5Google · Official$196.00
6OpenRouter$196.00
Switch to QiHang to save $160.50/mo (82%)
Note: this is a gateway; verify availability and rate limits yourself.
Benchmark4 items
| Name | Conditions | Score | Metric | Source |
|---|---|---|---|---|
| SWE-Bench Pro | dataset: public | 34.63 | resolve rate | Source ↗ |
| SWE-Atlas Codebase QnA | harness: Mini-SWE-Agent | 8.2 | score | Source ↗ |
| SWE-Atlas Refactoring | harness: Mini-SWE-Agent | 10 | score | Source ↗ |
| SWE-Atlas Test Writing | harness: Mini-SWE-Agent | 30.3 | score | Source ↗ |
The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.
Reasoning control
effort = loweffort = mediumeffort = higheffort = minimalToggle (on / off)effort = xhigheffort = maxeffort = none
3 / 25 providers expose no reasoning control (reasoning_options: []).
Related models
Gemini 3.7 Flashsame series$0.75 / $3.75Gemini 3.7 Flash (Google Vertex AI)same series$0.75 / $3.75Gemini 3.7 Flash (Google AI Studio)same series$0.75 / $3.75DeepSeek V4 Procheaper alternative$0.435 / $0.87DeepSeek V4 Flashcheaper alternative$0.14 / $0.28MiniMax-M3cheaper alternative$0.30 / $1.20GPT-5.6 Lunacheaper alternative$0.20 / $1.20
Price historyone sample accumulated per data sync
Input listOutput listMin blended