← Model list
Qwen Flash (Alibaba Cloud)
alibaba·alibaba/qwen-flash-2·GA·Closed·qwen series
Efficient Qwen model for fast chat, extraction, and high-volume workloads
At a glance
Reference price$0.05 / $0.40 per 1M
Cache read $0.01 · LLM Gateway
Lowest paid$0.05 / $0.40
LLM Gateway Gateway
Context1,000,000
Output limit32,000
Capabilities
Reasoning✓ Tool useStructured output✓ TemperatureAttachments
Modalities
Text
Knowledge cutoff2024-04
Released / updated2025-07-28 / 2025-07-28
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 1 providers1 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| LLM Gateway alibaba/qwen-flash | Gateway | $0.05 | $0.40 | $0.01 | $0.063 | 1,000,000 | 32,000 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1LLM Gateway$25.20
The cheapest paid channel is the only channel.
Related models
Qwen3.5 0.8Bsame series$0.06 / $0.12Qwen3.5 4Bsame series$0.10 / $0.20Qwen3.8 27B TEEsame series$0.40 / $3.00Lyria 3 Clip Previewcheaper alternative$0 / $0Lyria 3 Pro Previewcheaper alternative$0 / $0
Price historyone sample accumulated per data sync
Input list $0.05Output list $0.40Min blended $0.138
Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-21). Each future sync adds a point, and once accumulated a line is drawn here.