LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

MiniMax-M2.5-highspeed

minimax·minimax/minimax-m2.5-highspeed·GA·Open weights·minimax series
High-speed MiniMax model for low-latency coding and agent workflows

MiniMax-M2.5-highspeed is offered by 12 providers on this page. Public prices are shown for 9 of them. Its official list price is $0.60 per 1M input tokens and $2.40 per 1M output tokens. It also has 2 covered by a paid subscription; free tiers usually carry rate limits, and subscription-covered access bills $0 per token only after the subscription fee.

The context window is 204,800 tokens, with an output limit of 131,072 tokens. It supports reasoning, tool use, and structured output. The weights are open, so it can also be self-hosted or served through a gateway of your choice. Providers report the capability flags inconsistently, so verify a specific feature against the host you plan to use. Its training knowledge cuts off at 2025-01-01.

Specs & pricing

Input / output per 1M tokens
Official price·MiniMax (minimax.cn)
$0.60 / $2.40
Blended $1.05 · Cache read $0.06
Lowest paid·NovitaAIGateway
$0.60 / $2.40
Blended $1.05 · 1× spread
Context
204,800
Output limit
131,072
Knowledge cutoff
2025-01-01
Released / updated
2026-02-13 / 2026-02-13
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
Text
Weights
Hugging Face

Available at 12 providers9 with public prices · 2 subscription-covered

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
MiniMax Token Plan (minimax.io)
MiniMax-M2.5-highspeed
GatewaySubscription$0$0204,800131,072
MiniMax Token Plan (minimax.cn)
MiniMax-M2.5-highspeed
GatewaySubscription$0$0204,800131,072
NovitaAIGateway$0.60$2.40$0.03—204,800131,072
MiniMax (minimax.cn)Official
MiniMax-M2.5-highspeed
First-party$0.60$2.40$0.06$0.38204,800131,072
Vercel AI GatewayCloud$0.60$2.40$0.03$0.38204,800131,000
OrcaRouterGateway$0.60$2.40$0.03$0.38204,800131,072
Merge GatewayGateway$0.60$2.40$0.06$0.38204,8008,192
MiniMax (minimax.io)Official
MiniMax-M2.5-highspeed
First-party$0.60$2.40$0.06$0.38204,800131,072
DevPass (LLM Gateway)Gateway$0.60$2.40$0.03$0.38204,800131,072
LLM GatewayGateway$0.60$2.40$0.03—204,800131,100
ZenMux
minimax/minimax-m2.5-lightning
Gateway$0.60$4.80$0.06$0.75204,800131,072
QiniuGateway————204,800128,000hostTable.undisclosed

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

budget_tokenseffort = loweffort = mediumeffort = higheffort = xhigheffort = maxToggle (on / off)

Interleaved thinking (reasoning between tool calls) is declared by 1 of 12 providers.

8 / 12 providers expose no reasoning control (reasoning_options: []).

Your usage cost

1NovitaAI$171.60
2Vercel AI Gateway$171.60
3OrcaRouter$171.60
4DevPass (LLM Gateway)$171.60
5LLM Gateway$171.60
6MiniMax (minimax.cn) · Official$175.20

2 more channels offer $0 (MiniMax Token Plan (minimax.io), MiniMax Token Plan (minimax.cn)); free tiers usually have rate limits and no SLA, excluded from ranking.

Switch to NovitaAI to save $3.60/mo (2%). The gap is small, so staying on the official channel is fine.
Note: this is a gateway; verify availability and rate limits yourself.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$2.40$02026-08-052026-09-032026-09-03 · Input list $0.602026-09-03 · Output list $2.402026-09-03 · Min blended $1.05

Related models

MiniMax-M3.1-Flash-Previewsame series$0 / $0MiniMax Latestsame series$0.30 / $1.20MiniMax H3 Maxsame series—GLM-5.3-Flashcheaper alternative$0.15 / $0.50DeepSeek V4.1 Flashcheaper alternative$0.15 / $0.60GPT-5.6 Lunacheaper alternative$0.20 / $1.20GPT-5.4 nanocheaper alternative$0.20 / $1.25
Data partly from models.dev (MIT) · MiniMax (minimax.cn) official docs ↗