LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Gemini 3 Flash

google·google/gemini-3-flash·GA·Closed·gemini-flash series
Fast Gemini model balancing multimodal reasoning, tool use, and cost

Gemini 3 Flash by google is offered by 3 providers on this page. Its reference price is $0.50 per 1M input tokens and $3.00 per 1M output tokens. The lowest paid channel is Poe at $0.40 / $2.40 per 1M, about 1.3× below the reference price. That channel is a third-party gateway, so confirm its availability and rate limits before depending on it.

The context window is 1,048,576 tokens at the reference host, but hosts report different limits, from 1,000,000 to 1,048,576, so the usable window depends on the provider you pick. It supports reasoning, tool use, and structured output. Accepted input modalities are Text, Image, Video, Audio, and PDF. Its training knowledge cuts off at 2025-03.

Specs & pricing

Input / output per 1M tokens
Reference price·OpenCode Zen
$0.50 / $3.00
Blended $1.13 · Cache read $0.05
Lowest paid·PoeGateway
$0.40 / $2.40
Blended $0.90 · 1.3× spread
Context
1,048,576
Output limit
65,536
Knowledge cutoff
2025-03
Released / updated
2025-12-17 / 2025-12-17
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImageVideoAudioPDF

Available at 3 providers3 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
PoeGateway$0.40$2.40$0.04—1,048,57665,536
Vercel AI GatewayCloud$0.50$3.00$0.05—1,000,000 ⚠65,000
OpenCode ZenGateway$0.50$3.00$0.05—1,048,57665,536

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = minimaleffort = loweffort = higheffort = medium

Your usage cost

1Poe$156.80
2Vercel AI Gateway$196.00
3OpenCode Zen$196.00
The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$3.00$02026-08-052026-08-132026-08-05 · Input list $0.502026-08-06 · Input list $0.502026-08-07 · Input list $0.502026-08-08 · Input list $0.502026-08-09 · Input list $0.502026-08-10 · Input list $0.502026-08-11 · Input list $0.502026-08-12 · Input list $0.502026-08-13 · Input list $0.502026-08-05 · Output list $3.002026-08-06 · Output list $3.002026-08-07 · Output list $3.002026-08-08 · Output list $3.002026-08-09 · Output list $3.002026-08-10 · Output list $3.002026-08-11 · Output list $3.002026-08-12 · Output list $3.002026-08-13 · Output list $3.002026-08-05 · Min blended $0.902026-08-06 · Min blended $0.902026-08-07 · Min blended $0.902026-08-08 · Min blended $0.902026-08-09 · Min blended $0.902026-08-10 · Min blended $0.902026-08-11 · Min blended $0.902026-08-12 · Min blended $0.902026-08-13 · Min blended $0.90

Related models

Gemini 3.8 Flashsame series$0.75 / $3.75Gemini 3.8 Flash (EU)same series$0.83 / $4.13Gemini 3.8 Flash (Vertex AI, US)same series$0.75 / $3.75GLM-5.3-Flashcheaper alternative$0.15 / $0.50DeepSeek V4.1 Flashcheaper alternative$0.15 / $0.60DeepSeek V4 Flash 0731cheaper alternative$0.45 / $1.34GPT-5.6 Lunacheaper alternative$0.20 / $1.20
Data partly from models.dev (MIT) · OpenCode Zen official docs ↗