LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Qwen3.8 Max Preview

alibaba·alibaba/qwen3.8-max-preview·Deprecated·Closed·qwen series·NEW
Preview Qwen flagship for million-token multimodal reasoning and long-horizon agentic workflows

Qwen3.8 Max Preview by alibaba is offered by 6 providers on this page. Public prices are shown for 4 of them. Its reference price is $2.50 per 1M input tokens and $7.50 per 1M output tokens. The lowest paid channel is AIHubMix at $0.338 / $1.01 per 1M, about 7.4× below the reference price. That channel is a third-party gateway, so confirm its availability and rate limits before depending on it. It also has 2 covered by a paid subscription; free tiers usually carry rate limits, and subscription-covered access bills $0 per token only after the subscription fee.

The context window is 1,000,000 tokens, with an output limit of 131,072 tokens. It supports reasoning, tool use, and structured output. Accepted input modalities are Text, Image, and Video.

Specs & pricing

Input / output per 1M tokens
Reference price·Impossibl
$2.50 / $7.50
Blended $3.75 · Cache read —
Lowest paid·AIHubMixGateway
$0.34 / $1.01
Blended $0.51 · 7.4× spread
Context
1,000,000
Output limit
131,072
Knowledge cutoff
—
Released / updated
2026-07-19 / 2026-07-19
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImageVideo

Available at 6 providers4 with public prices · 2 subscription-covered

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
Alibaba Token PlanGatewaySubscription$0$01,000,000131,072deprecated
Alibaba Token Plan (China)GatewaySubscription$0$01,000,000131,072deprecated
AIHubMixGateway$0.34$1.01$0.068$0.421,000,000131,072
Charm Hyper
qwen3.8-max
Gateway$2.00$6.00$0.25—1,000,00065,536
DevPass (LLM Gateway)
qwen3.8-max
Gateway$2.00$6.00$0.25$2.501,000,0001,000,000
ImpossiblGateway$2.50$7.50——1,000,000131,072

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = loweffort = mediumeffort = xhighbudget_tokensToggle (on / off)effort = noneeffort = minimaleffort = high

Interleaved thinking (reasoning between tool calls) is declared by 3 of 6 providers.

Your usage cost

1AIHubMix$85.85
2Charm Hyper$490.00
3DevPass (LLM Gateway)$490.00
4Impossibl$875.00

2 more channels offer $0 (Alibaba Token Plan, Alibaba Token Plan (China)); free tiers usually have rate limits and no SLA, excluded from ranking.

The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$7.50$02026-08-052026-09-282026-08-05 · Input list $2.502026-08-06 · Input list $2.502026-08-07 · Input list $2.502026-08-08 · Input list $2.502026-08-09 · Input list $2.502026-08-10 · Input list $2.502026-08-11 · Input list $2.502026-08-12 · Input list $2.502026-08-13 · Input list $2.502026-09-10 · Input list $2.502026-09-28 · Input list $2.502026-08-05 · Output list $7.502026-08-06 · Output list $7.502026-08-07 · Output list $7.502026-08-08 · Output list $7.502026-08-09 · Output list $7.502026-08-10 · Output list $7.502026-08-11 · Output list $7.502026-08-12 · Output list $7.502026-08-13 · Output list $7.502026-09-10 · Output list $7.502026-09-28 · Output list $7.502026-08-05 · Min blended $2.382026-08-06 · Min blended $2.382026-08-07 · Min blended $2.382026-08-08 · Min blended $2.382026-08-09 · Min blended $2.382026-08-10 · Min blended $2.382026-08-11 · Min blended $2.382026-08-12 · Min blended $2.382026-08-13 · Min blended $2.722026-09-10 · Min blended $3.002026-09-28 · Min blended $0.51

Benchmark15 items

NameConditionsScoreMetricSource
Terminal-Benchvariant: xhigh · v2.186.6accuracySource ↗
SWE-Bench Proharness: Claude Code · variant: xhigh67.7resolve rateSource ↗
DeepSWEharness: Claude Code · variant: xhigh · v1.156.6resolve rateSource ↗
NL2Repoharness: Claude Code · variant: xhigh55.9resolve rateSource ↗
FrontierSWEharness: Claude Code · variant: xhigh73.5dominance scoreSource ↗
MLS-Bench-Liteharness: Claude Code · variant: xhigh41scoreSource ↗
AutomationBenchvariant: xhigh · dataset: 600-task public subset27.3pass@1Source ↗
Toolathlon Verifiedvariant: xhigh72.5pass@1Source ↗
WideSearchvariant: xhigh81.9F1Source ↗
Humanity's Last Examvariant: xhigh, with tools56.2accuracySource ↗
GPQA Diamondvariant: xhigh92.6accuracySource ↗
Humanity's Last Examvariant: xhigh, no tools43.6accuracySource ↗
IFBenchvariant: xhigh82.8scoreSource ↗
OSWorld-Verifiedvariant: xhigh86.1success rateSource ↗
MMMU Provariant: xhigh82.3accuracySource ↗

The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.

Related models

Qwen-Image-2.1same series$0 / $0Qwen 3.8 27B Hemingwaysame series$0.25 / $1.50Qwen 3.8 27B Cybersecuritysame series$0.10 / $0.60GLM-5.2cheaper alternative$1.40 / $4.40GLM-5.3cheaper alternative$1.40 / $4.40GLM-5.3-Flashcheaper alternative$0.15 / $0.50DeepSeek V4.1 Flashcheaper alternative$0.15 / $0.60
Data partly from models.dev (MIT) · Impossibl official docs ↗