← Model list
Gemini 3.5 Flash Lite
google·google/gemini-3.5-flash-lite·GA·Closed·gemini-flash-lite series·NEW
Fast Gemini model balancing multimodal reasoning, tool use, and cost
At a glance
Official price$0.30 / $2.50 per 1M
Cache read $0.03 · Vertex
Lowest paid$0.27 / $2.25
Requesty Gateway · 1.4× spread
Context1,048,576
⚠ Providers report 1,000,000–1,048,576; the table below is authoritative
Output limit65,536
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImageAudioPDFVideo
Knowledge cutoff2026-03
Released / updated2026-07-21 / 2026-07-21
Quality & performanceArtificial Analysis · Intelligence Index v4.1
Intelligence37.4
Value49
Coding49.3
Agentic27.2
Output speed332 tok/s
TTFT9.76 s
Cost per task$0.0965
Value formulaIQ 37.4 ÷ min blended $0.765 = 49
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 20 providers20 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Requesty | Gateway | $0.27 | $2.25 | $0.027 | — | 1,048,576 | 65,535 | |
| NanoGPT | Gateway | $0.30 | $2.50 | $0.03 | $0.083 | 1,048,576 | 65,536 | |
| VertexOfficial | First-party | $0.30 | $2.50 | $0.03 | — | 1,048,576 | 65,536 | |
| GoogleOfficial | First-party | $0.30 | $2.50 | $0.03 | — | 1,048,576 | 65,536 | |
| Impossibl | Gateway | $0.30 | $2.50 | $0.03 | — | 1,048,576 | 65,536 | |
| OpenRouter | Gateway | $0.30 | $2.50 | $0.03 | $0.083 | 1,048,576 | 65,536 | |
| CrossModel | Gateway | $0.30 | $2.50 | $0.03 | $0.30 | 1,048,576 | 65,536 | |
| Merge Gateway | Gateway | $0.30 | $2.50 | $0.03 | — | 1,048,576 | 65,536 | |
| Ofox | Gateway | $0.30 | $2.50 | $0.03 | $0.083 | 1,048,576 | 65,536 | |
| Neon gemini-3-5-flash-lite | Cloud | $0.30 | $2.50 | $0.03 | — | 1,048,576 | 65,536 | |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Requesty$137.34
2NanoGPT$152.60
3Vertex · Official$152.60
4Google · Official$152.60
5Impossibl$152.60
6OpenRouter$152.60
Switch to Requesty to save $15.26/mo (10%). The gap is small, so staying on the official channel is fine.
Note: this is a gateway; verify availability and rate limits yourself.
Benchmark9 items
| Name | Conditions | Score | Metric | Source |
|---|---|---|---|---|
| SWE-Bench Pro | — | 54.2 | resolve rate | Source ↗ |
| Terminal-Bench | harness: Terminus 2 · v2.1 | 54 | accuracy | Source ↗ |
| MLE-Bench | — | 39.2 | average position score | Source ↗ |
| GDPval-AA | vv2 | 1140 | Elo | Source ↗ |
| OSWorld-Verified | — | 74 | success rate | Source ↗ |
| CharXiv Reasoning | variant: no tools | 74.5 | accuracy | Source ↗ |
| CharXiv Reasoning | variant: with tools | 76.5 | accuracy | Source ↗ |
| GDM-MRCR | variant: 128k average, 8-needle · vv2 | 72.2 | accuracy | Source ↗ |
| GDM-MRCR | variant: 1M pointwise, 8-needle · vv2 | 21.3 | accuracy | Source ↗ |
The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.
Reasoning control
effort = noneeffort = loweffort = mediumeffort = higheffort = maxbudget_tokenseffort = minimalToggle (on / off)
2 / 20 providers expose no reasoning control (reasoning_options: []).
Related models
Gemini 3.5 Flash Lite (Google Vertex AI)same series$0.30 / $2.50Gemini 3.5 Flash Lite (Google AI Studio)same series$0.30 / $2.50Gemini 3.5 Flash Lite (EU)same series$0.30 / $2.50DeepSeek V4 Flashcheaper alternative$0.14 / $0.28GPT-5.6 Lunacheaper alternative$0.20 / $1.20MiMo-V2.5cheaper alternative$0.14 / $0.28Gemini 2.5 Flash-Litecheaper alternative$0.10 / $0.40
Price historyone sample accumulated per data sync
Input listOutput listMin blended