LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Qwen3.8 Flash

alibaba·alibaba/qwen3.8-flash·GA·Closed·qwen series·NEW

Qwen vision-language model for visual reasoning, documents, and agent tasks

At a glance
Reference price$0.16 / $0.47 per 1M
Cache read $0.016 · Vercel AI Gateway
Lowest paid$0.15 / $0.47
OpenRouter Gateway · 1.1× spread
Context991,000
⚠ Providers report 991,000–1,000,000; the table below is authoritative
Output limit128,000
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImageVideoPDF
Knowledge cutoff—
Released / updated2026-08-26 / 2026-08-26

Quality & performance

Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.

Available at 6 providers6 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
OpenRouterGateway$0.15$0.47$0.016$0.201,000,000 ⚠131,072
Kilo GatewayGateway$0.15$0.47$0.016$0.201,000,000 ⚠131,072
NanoGPTGateway$0.16$0.47$0.016$0.20991,808 ⚠131,072
Charm HyperGateway$0.16$0.47$0.016—1,000,000 ⚠128,000
EmpirioLabs AI
qwen3-8-flash
Gateway$0.16$0.47$0.16—1,000,000 ⚠131,072
Vercel AI GatewayCloud$0.16$0.47$0.016$0.20991,000128,000

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1OpenRouter$37.42
2Kilo Gateway$37.42
3NanoGPT$38.22
4Charm Hyper$38.22
5Vercel AI Gateway$38.22
6EmpirioLabs AI$55.50
The cheapest paid channel is the only channel.

Benchmark

No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.

Reasoning control

Toggle (on / off)budget_tokenseffort = noneeffort = higheffort = loweffort = mediumeffort = max

1 / 6 providers expose no reasoning control (reasoning_options: []).

Related models

Qwen 3.8 27B Uncensored Thinkingsame series$0.18 / $0.50Qwen 3.8 27B Obliteratedsame series$0.18 / $0.50Qwen 3.8 27B Obliterated Thinkingsame series$0.18 / $0.50GLM-5.3-Flashcheaper alternative$0.075 / $0.25Qwen3.7 Flashcheaper alternative$0.03 / $0.118Qwen Flashcheaper alternative$0.022 / $0.216Qwen Turbocheaper alternative$0.044 / $0.087

Price historyone sample accumulated per data sync

Input list $0.16Output list $0.47Min blended $0.23

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-27). Each future sync adds a point, and once accumulated a line is drawn here.

Data from models.dev (MIT) · Vercel AI Gateway official docs ↗