LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Umans Flash

alibaba·alibaba/umans-flash·GA·Open weights·qwen series
Open multimodal Qwen MoE for local agents that need vision, audio, and code

Umans Flash by alibaba is offered by 2 providers on this page. Public prices are shown for 1 of them. Its reference price is $0.15 per 1M input tokens and $1.00 per 1M output tokens. It also has 1 covered by a paid subscription; free tiers usually carry rate limits, and subscription-covered access bills $0 per token only after the subscription fee.

The context window is 262,144 tokens, with an output limit of 32,768 tokens. It supports reasoning, tool use, and structured output. Accepted input modalities are Text and Image. The weights are open, so it can also be self-hosted or served through a gateway of your choice.

Specs & pricing

Input / output per 1M tokens
Reference price·Umans AI
$0.15 / $1.00
Blended $0.362 · Cache read $0.05
Lowest paid·Umans AIGateway
$0.15 / $1.00
Blended $0.362
1 more $0 channels
Context
262,144
Output limit
32,768
Knowledge cutoff
—
Released / updated
2026-04-17 / 2026-04-17
Capabilities
✓ Reasoning✓ Tool use✓ Structured outputTemperature✓ Attachments
Modalities
TextImage

Available at 2 providers1 with public prices · 1 subscription-covered

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
Umans AI Coding PlanGatewaySubscription$0$0262,144262,144
Umans AIGateway$0.15$1.00$0.05—262,14432,768

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Umans AI$68.00

1 more channels offer $0 (Umans AI Coding Plan); free tiers usually have rate limits and no SLA, excluded from ranking.

The cheapest paid channel is the only channel.

Reasoning control

Toggle (on / off)effort = loweffort = mediumeffort = high

Related models

Qwen3.8 Max 0902same series$2.00 / $6.00Qwen 3.8 27B Fablesame series$0.25 / $1.50Qwen3.8 Flash Nextsame series$0.201 / $0.50DeepSeek V4 Flashcheaper alternative$0.14 / $0.28GLM-5.3-Flashcheaper alternative$0.075 / $0.25GPT-5 Nanocheaper alternative$0.05 / $0.40MiMo-V2.5cheaper alternative$0.14 / $0.28

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$1.00$02026-08-052026-08-132026-08-05 · Input list $0.152026-08-06 · Input list $0.152026-08-07 · Input list $0.152026-08-08 · Input list $0.152026-08-09 · Input list $0.152026-08-10 · Input list $0.152026-08-11 · Input list $0.152026-08-12 · Input list $0.152026-08-13 · Input list $0.152026-08-05 · Output list $1.002026-08-06 · Output list $1.002026-08-07 · Output list $1.002026-08-08 · Output list $1.002026-08-09 · Output list $1.002026-08-10 · Output list $1.002026-08-11 · Output list $1.002026-08-12 · Output list $1.002026-08-13 · Output list $1.002026-08-05 · Min blended $0.3622026-08-06 · Min blended $0.3622026-08-07 · Min blended $0.3622026-08-08 · Min blended $0.3622026-08-09 · Min blended $0.3622026-08-10 · Min blended $0.3622026-08-11 · Min blended $0.3622026-08-12 · Min blended $0.3622026-08-13 · Min blended $0.362
Data partly from models.dev (MIT) · Umans AI official docs ↗