← Model list
gemini-2.5-flash-lite-preview-09-2025
google·google/gemini-2.5-flash-lite-preview-09-2025·GA·Closed·gemini series
Low-latency Gemini model for high-volume multimodal and agent workloads
At a glance
Reference price$0.10 / $0.40 per 1M
Cache read — · 302.AI
Lowest paid$0.09 / $0.36
Jiekou.AI Gateway · 1.1× spread
Context1,000,000
⚠ Providers report 1,000,000–1,048,756; the table below is authoritative
Output limit65,536
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImageAudioPDFVideo
Knowledge cutoff2025-01
Released / updated2025-09-25 / 2026-01
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 3 providers3 with public prices
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1NanoGPT$29.20
2Jiekou.AI$36.00
3302.AI$40.00
The cheapest paid channel is the only channel.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
effort = noneeffort = loweffort = mediumeffort = high
Related models
Gemini Omni Flash Previewsame series$1.50 / $17.50Gemini Embedding 2same series$0.20 / $0Gemini Robotics-ER 1.6 Previewsame series$1.00 / $5.00Qwen3.7 Flashcheaper alternative$0.03 / $0.118Qwen Flashcheaper alternative$0.022 / $0.216Qwen Turbocheaper alternative$0.044 / $0.087Laguna S 2.1cheaper alternative$0 / $0
Price historyone sample accumulated per data sync
Input listOutput listMin blended