← Model list
Gemini 3.1 Flash Lite
google·google/gemini-3.1-flash-lite·GA·Closed·gemini-flash-lite series
Low-latency Gemini model for high-volume multimodal and agent workloads
At a glance
Official price$0.25 / $1.50 per 1M
Cache read $0.025 · Vertex
Lowest paid$0.225 / $1.35
Requesty Gateway · 1.2× spread
1 more $0 channels
Context1,048,576
⚠ Providers report 1,000,000–1,048,576; the table below is authoritative
Output limit65,536
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImageAudioPDFVideo
Knowledge cutoff2025-01
Released / updated2026-05-07 / 2026-05-07
Quality & performanceArtificial Analysis · Intelligence Index v4.1
Intelligence25.6
Value51
Coding34.7
Agentic6.5
Output speed303 tok/s
TTFT5.56 s
Cost per task$0.0432
Value formulaIQ 25.6 ÷ min blended $0.506 = 51
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 21 providers20 with public prices · 1 free
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Kenari gemini-3-1-flash-lite | Gateway | Free | — | — | 1,048,576 | 65,536 | ||
| Requesty | Gateway | $0.225 tiered >200K: $0.45 | $1.35 | $0.022 | $0.075 | 1,048,576 | 65,535 | |
| NanoGPT | Gateway | $0.25 | $1.50 | $0.025 | $0.083 | 1,048,576 | 65,536 | |
| VertexOfficial | First-party | $0.25 | $1.50 | $0.025 | — | 1,048,576 | 65,536 | |
| GoogleOfficial | First-party | $0.25 | $1.50 | $0.025 | — | 1,048,576 | 65,536 | |
| Impossibl | Gateway | $0.25 | $1.50 | $0.025 | — | 1,048,576 | 65,536 | |
| OpenRouter | Gateway | $0.25 | $1.50 | $0.025 | $0.083 | 1,048,576 | 65,536 | |
| NEAR AI Cloud | Gateway | $0.25 | $1.50 | $0.025 | — | 1,048,576 | 65,536 | |
| Merge Gateway | Gateway | $0.25 | $1.50 | $0.025 | — | 1,048,576 | 65,536 | |
| ZenMux | Gateway | $0.25 | $1.50 | $0.025 | — | 1,048,576 | 65,536 | |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Requesty$88.20
2NanoGPT$98.00
3Vertex · Official$98.00
4Google · Official$98.00
5Impossibl$98.00
6OpenRouter$98.00
1 more channels offer $0 (Kenari); free tiers usually have rate limits and no SLA, excluded from ranking.
Switch to Requesty to save $9.80/mo (10%). The gap is small, so staying on the official channel is fine.
Note: this is a gateway; verify availability and rate limits yourself.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
effort = minimaleffort = loweffort = mediumeffort = higheffort = noneeffort = maxbudget_tokensToggle (on / off)
3 / 21 providers expose no reasoning control (reasoning_options: []).
Related models
Gemini 3.5 Flash Litesame series$0.30 / $2.50Gemini 3.5 Flash Lite (Google Vertex AI)same series$0.30 / $2.50Gemini 3.5 Flash Lite (Google AI Studio)same series$0.30 / $2.50DeepSeek V4 Flashcheaper alternative$0.14 / $0.28MiMo-V2.5cheaper alternative$0.14 / $0.28Gemini 2.5 Flash-Litecheaper alternative$0.10 / $0.40Qwen Pluscheaper alternative$0.115 / $0.287
Price historyone sample accumulated per data sync
Input listOutput listMin blended