LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Qwen3-Coder-Next-FP8-no-thinking

alibaba·alibaba/qwen3-coder-next-fp8-no-thinking·GA·Open weights·qwen series
Qwen coding model for software agents, repository edits, and code reasoning

Specs & pricing

Input / output per 1M tokens
Reference price·InferX
$0 / $0
Blended $0 · Cache read —
Lowest paid
No paid channels
1 more $0 channels
Context
260,000
Output limit
65,536
Knowledge cutoff
2025-04
Released / updated
2026-02-03 / 2026-02-03
Capabilities
Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
Modalities
Text

Available at 1 providers0 with public prices · 1 free

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
InferX
Qwen3-Coder-Next-FP8-no-thinking
GatewayFree——260,00065,536

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

This model is $0 across all listed channels (InferX).

Price historyone sample accumulated per data sync

Input list $0Output list $0Min blended —

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-17). Each future sync adds a point, and once accumulated a line is drawn here.

Related models

Qwen-Image-2.1same series$0 / $0Qwen 3.8 27B Hemingwaysame series$0.25 / $1.50Qwen 3.8 27B Cybersecuritysame series$0.10 / $0.60
Data partly from models.dev (MIT) · InferX official docs ↗