LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Step TTS 2

stepfun·stepfun/step-tts-2·GA·Closed·step series·Speech / Audio
Speech generation model for controllable voice, narration, and audio delivery

Specs & pricing

Input / output per 1M tokens
Official price
Price undisclosed
Lowest paid
No paid channels
Context
—
Output limit
—
Knowledge cutoff
—
Released / updated
2026-03-01 / 2026-07-02
Capabilities
ReasoningTool use? Structured outputTemperatureAttachments
Modalities
Text

Available at 2 providers0 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
StepFun (Global)OfficialFirst-party——————hostTable.undisclosed
StepFun (China)OfficialFirst-party——————hostTable.undisclosed

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

This model has no public prices; can't estimate.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.0001$02026-08-052026-08-13

Artificial Analysis arenas1 arenas

Text to speech1134±12#26 of 93

Elo from human preference matchups run by Artificial Analysis. Each arena is anchored separately, so Elo is only comparable within one arena, against the other models on that board, and the 95% confidence interval (shown as ±) indicates how much of a gap is noise. Rank is over the whole arena, including models not listed here.

Related models

StepAudio 2.5 ASRsame series—StepAudio 2.5 TTSsame series—
Data partly from models.dev (MIT)