← Model list

Qwen Flash (Alibaba Cloud)

alibaba·alibaba/qwen-flash-2·GA·Closed·qwen series

Efficient Qwen model for fast chat, extraction, and high-volume workloads

At a glance
Reference price$0.05 / $0.40 per 1M
Cache read $0.01 · LLM Gateway
Lowest paid$0.05 / $0.40
LLM Gateway Gateway
Context1,000,000
Output limit32,000
Capabilities
ReasoningTool useStructured outputTemperatureAttachments
Modalities
Text
Knowledge cutoff2024-04
Released / updated2025-07-28 / 2025-07-28

Quality & performance

Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
LLM Gateway
alibaba/qwen-flash
Gateway$0.05$0.40$0.01$0.0631,000,00032,000

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1LLM Gateway$25.20
The cheapest paid channel is the only channel.

Related models

Price historyone sample accumulated per data sync

Input list $0.05Output list $0.40Min blended $0.138

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-21). Each future sync adds a point, and once accumulated a line is drawn here.