Open-weight experimental preview of the Qwen4 architecture: hybrid-attention MoE (125B total, 6B active) with vision encoder for coding, agent tasks, and image and video understanding
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-29). Each future sync adds a point, and once accumulated a line is drawn here.