LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Kimi Fast Latest

moonshotai·moonshotai/kimi-fast-latest·GA·Open weights·kimi-k3 series·NEW
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work

Specs & pricing

Input / output per 1M tokens
Reference price·Fireworks AI
$4.50 / $22.50
Blended $9.00 · Cache read $0.45
Lowest paid·Fireworks AICloud
$4.50 / $22.50
Blended $9.00
Context
1,048,576
Output limit
131,072
Knowledge cutoff
—
Released / updated
2026-07-27 / 2026-09-15
Capabilities
✓ Reasoning✓ Tool use✓ Structured outputTemperature✓ Attachments
Modalities
TextImage

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
Fireworks AICloud$4.50$22.50$0.45—1,048,576131,072

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

Toggle (on / off)effort = loweffort = mediumeffort = higheffort = maxbudget_tokens ≥ 1,024

Interleaved thinking (reasoning between tool calls) is declared by 1 of 1 providers.

Your usage cost

1Fireworks AI$1,539.00
The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input list $4.50Output list $22.50Min blended $9.00

Price history accumulates from each data sync; currently only 1 sample(s) (2026-09-16). Each future sync adds a point, and once accumulated a line is drawn here.

Related models

Kimi K3 Fastsame series$4.50 / $22.50Kimi K3same series$3.00 / $15.00Kimi K3-256Ksame series$0 / $0GLM-5.2cheaper alternative$1.40 / $4.40GLM-5.3cheaper alternative$1.40 / $4.40GLM-5.3-Flashcheaper alternative$0.15 / $0.50DeepSeek V4.1 Flashcheaper alternative$0.15 / $0.60
Data partly from models.dev (MIT) · Fireworks AI official docs ↗