LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Qwen3.8 Flash Next

alibaba·alibaba/qwen3.8-flash-next·GA·Open weights·qwen series·NEW

Open-weight experimental preview of the Qwen4 architecture: hybrid-attention MoE (125B total, 6B active) with vision encoder for coding, agent tasks, and image and video understanding

At a glance
Reference price$0.201 / $0.50 per 1M
Cache read $0.05 · Cortecs
Lowest paid$0.15 / $0.47
AMD Gateway · 1.3× spread
Context262,144
Output limit64,000
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImage
Knowledge cutoff—
Released / updated2026-08-27 / 2026-08-27

Quality & performanceArtificial Analysis · Intelligence Index v4.1

Intelligence55.8
Value243
Coding73.1
Agentic56.4
Output speed73 tok/s
TTFT3.14 s
Cost per task$0.0959
Value formulaIQ 55.8 ÷ min blended $0.230 = 243

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 2 providers2 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
AMD
Qwen3.8-Flash-Next
Gateway$0.15$0.47$0.016—262,144131,072
CortecsGateway$0.201$0.50$0.05—262,14464,000

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1AMD$37.42
2Cortecs$47.08
The cheapest paid channel is the only channel.

Benchmark

No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.

Reasoning control

effort = loweffort = mediumeffort = xhigh

Related models

Qwen 3.8 27B Fablesame series$0.20 / $1.40Qwen3.8 Flashsame series$0.119 / $0.401Qwen3.8 Flash (NovitaAI)same series$0.15 / $0.47GLM-5.3-Flashcheaper alternative$0.075 / $0.25GPT-5 Nanocheaper alternative$0.05 / $0.40GLM-4.7-Flashcheaper alternative$0 / $0Hy3cheaper alternative$0 / $0

Price historyone sample accumulated per data sync

Input list $0.201Output list $0.50Min blended $0.23

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-29). Each future sync adds a point, and once accumulated a line is drawn here.

Data from models.dev (MIT) · Cortecs official docs ↗