← Model list
Qwen3 VL Flash
alibaba·alibaba/qwen3-vl-flash-2·GA·Closed·qwen series
Qwen vision-language model for visual reasoning, documents, and agent tasks
At a glance
Reference price$0.05 / $0.40 per 1M
Cache read $0.01 · DevPass (LLM Gateway)
Lowest paid$0.05 / $0.40
DevPass (LLM Gateway) Gateway
Context262,144
Output limit32,000
Capabilities
Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImage
Knowledge cutoff—
Released / updated2025-10-09 / 2025-10-09
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 1 providers1 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| DevPass (LLM Gateway) qwen3-vl-flash | Gateway | $0.05 | $0.40 | $0.01 | — | 262,144 | 32,000 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1DevPass (LLM Gateway)$25.20
The cheapest paid channel is the only channel.
Related models
Qwen3.5 0.8Bsame series$0.06 / $0.12Qwen3.5 4Bsame series$0.10 / $0.20Qwen3.8 27B TEEsame series$0.40 / $3.00Lyria 3 Clip Previewcheaper alternative$0 / $0Granite 4.1 8Bcheaper alternative$0.05 / $0.10Lyria 3 Pro Previewcheaper alternative$0 / $0Google Gemma 3 12Bcheaper alternative$0.05 / $0.15
Price historyone sample accumulated per data sync
Input list $0.05Output list $0.40Min blended $0.138
Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-21). Each future sync adds a point, and once accumulated a line is drawn here.