← Model list
Gemini 2.5 Flash Preview (09/2025) – Thinking
google·google/gemini-2.5-flash-preview-09-2025-thinking·GA·Closed·gemini series
Compact GPT model for low-latency assistance and high-volume workloads
At a glance
Reference price$0.30 / $2.50 per 1M
Cache read $0.03 · NanoGPT
Lowest paid$0.30 / $2.50
NanoGPT Gateway
Context1,048,756
Output limit65,536
Capabilities
✓ Reasoning✓ Tool use✓ Structured output? Temperature✓ Attachments
Modalities
TextImageAudioPDF
Knowledge cutoff—
Released / updated2025-09-25 / 2025-09-25
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 1 providers1 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NanoGPT | Gateway | $0.30 | $2.50 | $0.03 | — | 1,048,756 | 65,536 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1NanoGPT$152.60
The cheapest paid channel is the only channel.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
effort = loweffort = mediumeffort = high
Related models
Gemini Omni Flash Previewsame series$1.50 / $17.50Gemini Embedding 2same series$0.20 / $0Gemini Robotics-ER 1.6 Previewsame series$1.00 / $5.00DeepSeek V4 Flashcheaper alternative$0.14 / $0.28GPT-5.6 Lunacheaper alternative$0.20 / $1.20MiMo-V2.5cheaper alternative$0.14 / $0.28Gemini 2.5 Flash-Litecheaper alternative$0.10 / $0.40
Price historyone sample accumulated per data sync
Input listOutput listMin blended