← Model list
Gemini 3.7 Flash
google·google/gemini-3.7-flash·GA·Closed·gemini-flash series·NEW
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
At a glance
Official price$0.75 / $3.75 per 1M
Cache read $0.075 · Vertex
Lowest paid$0.375 / $1.88
NanoGPT Gateway · 5× spread
Context1,048,576
⚠ Providers report 1,000,000–1,048,576; the table below is authoritative
Output limit65,536
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImageAudioPDFVideo
Knowledge cutoff2026-03
Released / updated2026-08-13 / 2026-08-13
Quality & performanceArtificial Analysis · Intelligence Index v4.1 · rep. tier high
Intelligence56
Value75
Coding76.1
Agentic45.1
Output speed323 tok/s
TTFT15.15 s
Cost per task$0.4022
Value formulaIQ 56 ÷ min blended $0.750 = 75
Reasoning tier → intelligence / speed (higher tier = stronger but slower)
lowIQ 50.9 · 303 tok/s
mediumIQ 53.4 · 317 tok/s
highIQ 56 · 323 tok/s
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 21 providers21 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NanoGPT | Gateway | $0.375 | $1.88 | $0.037 | $0.021 | 1,048,576 | 65,536 | |
| OpenRouter | Gateway | $0.375 | $1.88 | $0.037 | $0.021 | 1,048,576 | 65,536 | |
| Requesty | Gateway | $0.60 | $3.00 | $0.06 | — | 1,048,576 | 65,535 | |
| Cortecs | Gateway | $0.75 | $3.75 | $0.075 | $0.038 | 1,048,576 | 1,048,576 | |
| VertexOfficial | First-party | $0.75 | $3.75 | $0.075 | — | 1,048,576 | 65,536 | |
| GoogleOfficial | First-party | $0.75 | $3.75 | $0.075 | — | 1,048,576 | 65,536 | |
| CrossModel | Gateway | $0.75 | $3.75 | $0.075 | $0.75 | 1,048,576 | 65,536 | |
| Merge Gateway | Gateway | $0.75 | $3.75 | $0.075 | — | 1,048,576 | 65,536 | |
| AIHubMix | Gateway | $0.75 | $3.75 | $0.075 | — | 1,048,576 | 65,536 | |
| Vercel AI Gateway | Cloud | $0.75 | $3.75 | $0.075 | — | 1,000,000 ⚠ | 65,536 | |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1NanoGPT$128.25
2OpenRouter$128.25
3Requesty$205.20
4Cortecs$256.50
5Vertex · Official$256.50
6Google · Official$256.50
Switch to NanoGPT to save $128.25/mo (50%)
Note: this is a gateway; verify availability and rate limits yourself.
Benchmark6 items
| Name | Conditions | Score | Metric | Source |
|---|---|---|---|---|
| FrontierCode | v1.1 Main | 43.6 | score | Source ↗ |
| DeepSWE | v1.1 | 65.3 | resolve rate | Source ↗ |
| Terminal-Bench | v2.1 | 85.8 | accuracy | Source ↗ |
| AutomationBench | dataset: private set | 30.4 | accuracy | Source ↗ |
| GDP.pdf | — | 34 | accuracy | Source ↗ |
| GDM-MRCR | variant: 128k average, 8-needle · vv2 | 97 | accuracy | Source ↗ |
The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.
Reasoning control
effort = minimaleffort = loweffort = mediumeffort = higheffort = noneeffort = maxbudget_tokens
3 / 21 providers expose no reasoning control (reasoning_options: []).
Related models
Gemini 3.7 Flash (Google Vertex AI)same series$0.75 / $3.75Gemini 3.7 Flash (Google AI Studio)same series$0.75 / $3.75Gemini 3.7 Flash (EU)same series$1.50 / $7.50DeepSeek V4 Procheaper alternative$0.435 / $0.87DeepSeek V4 Flashcheaper alternative$0.14 / $0.28DeepSeek V4 Flash 0731cheaper alternative$0.479 / $1.44MiniMax-M3cheaper alternative$0.30 / $1.20
Price historyone sample accumulated per data sync
Input list $0.75Output list $3.75Min blended $0.75
Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-14). Each future sync adds a point, and once accumulated a line is drawn here.