LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Qwen3 30B A3B 2507

alibaba·alibaba/qwen3-30b-a3b-2507·GA·Open weights·qwen series
Qwen instruction model for multilingual chat, reasoning, and tool use

Qwen3 30B A3B 2507 by alibaba is currently listed from a single provider. Its reference price is $0 per 1M tokens. It also has 1 free ($0) channel; free tiers usually carry rate limits, and subscription-covered access bills $0 per token only after the subscription fee.

Artificial Analysis rates it 9.8 on the Intelligence Index, with 12.1 for coding and 0.9 for agentic tasks. Median output speed is 145 tokens per second, with 2.37s to the first token. Running one task of the Artificial Analysis suite costs about $0.0766, which reflects how many tokens its reasoning consumes rather than the unit price alone.

The context window is 262,144 tokens, with an output limit of 16,384 tokens. It supports tool use. The weights are open, so it can also be self-hosted or served through a gateway of your choice. Its training knowledge cuts off at 2025-04.

Specs & pricing

Input / output per 1M tokens
Reference price·LMStudio
$0 / $0
Blended $0 · Cache read —
Lowest paid
No paid channels
1 more $0 channels
Context
262,144
Output limit
16,384
Knowledge cutoff
2025-04
Released / updated
2025-07-30 / 2025-07-30
Capabilities
Reasoning✓ Tool use? Structured output✓ TemperatureAttachments
Modalities
Text

Quality & performanceArtificial Analysis · Intelligence Index v4.3 · rep. tier Reasoning

Intelligence9.8
Value—
Coding12.1
Agentic0.9
Output speed145 tok/s
TTFT2.37 s
Cost per task$0.077
Reasoning tier → intelligence / speed (higher tier = stronger but slower)
defaultIQ 7.5 · 145 tok/s
ReasoningIQ 9.8 · 145 tok/s

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 1 providers0 with public prices · 1 free

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
LMStudioCloudFree——262,14416,384

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

This model is $0 across all listed channels (LMStudio).

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.0001$02026-08-052026-08-132026-08-05 · Input list $02026-08-06 · Input list $02026-08-07 · Input list $02026-08-08 · Input list $02026-08-09 · Input list $02026-08-10 · Input list $02026-08-11 · Input list $02026-08-12 · Input list $02026-08-13 · Input list $02026-08-05 · Output list $02026-08-06 · Output list $02026-08-07 · Output list $02026-08-08 · Output list $02026-08-09 · Output list $02026-08-10 · Output list $02026-08-11 · Output list $02026-08-12 · Output list $02026-08-13 · Output list $0

Artificial Analysis evaluations14 items

GPQA Diamond70.7%
Humanity's Last Exam10.3%
MMLU-Pro80.5%
SciCode33.0%
LiveCodeBench70.7%
Terminal-Bench Hard5.3%
Terminal-Bench 2.11.5%
AIME 202556.3%
AIME90.7%
MATH-50097.6%
τ²-Bench Telecom28.1%
τ³-Bench Banking5.4%
AA-LCR61.3%
IFBench50.7%

Individual evaluations run by Artificial Analysis, on the same reasoning tier as the intelligence score above. Each benchmark has its own task set and harness, so rows are not comparable with one another. The Intelligence Index above draws on a different, newer set of evaluations.

Related models

Qwen-Image-2.1same series$0 / $0Qwen 3.8 27B Hemingwaysame series$0.25 / $1.50Qwen 3.8 27B Cybersecuritysame series$0.10 / $0.60
Data partly from models.dev (MIT) · LMStudio official docs ↗