LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Qwen3.8 2.4T A95B (Max)

alibaba·alibaba/qwen3.8-2.4t-a95b-2·GA·Open weights·qwen series·NEW

Open-weight sparse MoE (2.4T total, 95B active), the open-weight twin of Qwen3.8 Max for coding, research, complex reasoning, and agentic workflows

At a glance
Reference price$2.00 / $6.00 per 1M
Cache read $0.25 · NanoGPT
Lowest paid$2.00 / $6.00
NanoGPT Gateway
Context991,000
Output limit65,536
Capabilities
Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImageVideoPDF
Knowledge cutoff—
Released / updated2026-08-12 / 2026-08-12

Quality & performance

Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NanoGPT
qwen/qwen3.8-2.4t-a95b
Gateway$2.00$6.00$0.25$2.50991,00065,536

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1NanoGPT$490.00
The cheapest paid channel is the only channel.

Related models

Qwen3.8 Flashsame series$0.16 / $0.47Qwen 3.8 27B Uncensored Thinkingsame series$0.18 / $0.50Qwen 3.8 27B Obliteratedsame series$0.18 / $0.50Qwen3 Coder Flashcheaper alternative$0.144 / $0.574Grok 4.20 (Non-Reasoning)cheaper alternative$1.25 / $2.50Llama 4 Maverick 17B Instructcheaper alternative$0.50 / $1.50Lyria 3 Clip Previewcheaper alternative$0 / $0

Price historyone sample accumulated per data sync

Input list $2.00Output list $6.00Min blended $3.00

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-13). Each future sync adds a point, and once accumulated a line is drawn here.

Data from models.dev (MIT) · NanoGPT official docs ↗