Fast Gemini model balancing multimodal reasoning, tool use, and cost
Gemini 3.8 Flash by google is offered by 2 providers on this page. Its reference price is $0.75 per 1M input tokens and $3.75 per 1M output tokens.
The context window is 1,048,576 tokens, with an output limit of 65,536 tokens. It supports reasoning, tool use, and structured output. Accepted input modalities are Text, Image, and Audio.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| DevPass (LLM Gateway) | Gateway | $0.75 | $3.75 | $0.075 | $0.083 | 1,048,576 | 1,048,576 | |
| LLM Gateway | Gateway | $0.75 | $3.75 | $0.075 | $0.083 | 1,048,576 | 65,536 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Price history accumulates from each data sync; currently only 1 sample(s) (2026-09-02). Each future sync adds a point, and once accumulated a line is drawn here.