← Model list

Qwen3.7 Flash

alibaba·alibaba/qwen3.7-flash·GA·Closed·qwen series·NEW

Lightweight multimodal Qwen model for high-throughput text, image, and video tasks

At a glance
Official price$0.03 / $0.118 per 1M
Cache read $0.003 · Alibaba (China)
Lowest paid$0.03 / $0.118
Alibaba (China) First-party · 6.8× spread
Context1,000,000
⚠ Providers report 991,000–1,000,000; the table below is authoritative
Output limit65,536
Capabilities
ReasoningTool useStructured outputTemperatureAttachments
Modalities
TextImageVideoPDF
Knowledge cutoff
Released / updated2026-07-15 / 2026-07-15

Quality & performance

Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.

Available at 9 providers9 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
Alibaba (China)OfficialFirst-party$0.03
tiered >32K: $0.089 · >256K: $0.178
$0.118$0.003$0.0371,000,00065,536
NanoGPTGateway$0.03$0.13$0.006$0.038991,80865,536
OpenRouterGateway$0.03
tiered >32K: $0.10 · >256K: $0.20
$0.13$0.006$0.0381,000,00065,536
EmpirioLabs AI
qwen3-7-flash
Gateway$0.03
tiered >32K: $0.10 · >256K: $0.20
$0.13$0.0061,000,00065,536
Vercel AI GatewayCloud$0.03$0.13$0.006$0.038991,00064,000
DevPass (LLM Gateway)Gateway$0.03$0.13$0.006$0.0371,000,0001,000,000
Kilo GatewayGateway$0.03$0.13$0.006$0.0381,000,00065,536
CrossModelGateway$0.04
tiered >32K: $0.10 · >256K: $0.19
$0.13$0.01$0.041,000,00065,536
Charm HyperGateway$0.20$0.80$0.041,000,00064,000

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Alibaba (China) · Official$8.65
2NanoGPT$9.62
3OpenRouter$9.62
4EmpirioLabs AI$9.62
5Vercel AI Gateway$9.62
6DevPass (LLM Gateway)$9.62
The cheapest paid channel is official.

Benchmark

No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.

Reasoning control

Toggle (on / off)budget_tokens ≥ 1 ≤ 131,072effort = noneeffort = higheffort = loweffort = mediumeffort = maxeffort = minimaleffort = xhigh

2 / 9 providers expose no reasoning control (reasoning_options: []).

Related models

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.118$02026-08-052026-08-132026-08-05 · Input list $0.032026-08-06 · Input list $0.032026-08-07 · Input list $0.032026-08-08 · Input list $0.032026-08-09 · Input list $0.032026-08-10 · Input list $0.032026-08-11 · Input list $0.032026-08-12 · Input list $0.032026-08-13 · Input list $0.032026-08-05 · Output list $0.1182026-08-06 · Output list $0.1182026-08-07 · Output list $0.1182026-08-08 · Output list $0.1182026-08-09 · Output list $0.1182026-08-10 · Output list $0.1182026-08-11 · Output list $0.1182026-08-12 · Output list $0.1182026-08-13 · Output list $0.1182026-08-05 · Min blended $0.0522026-08-06 · Min blended $0.0522026-08-07 · Min blended $0.0522026-08-08 · Min blended $0.0522026-08-09 · Min blended $0.0522026-08-10 · Min blended $0.0522026-08-11 · Min blended $0.0522026-08-12 · Min blended $0.0522026-08-13 · Min blended $0.052