← Model list

Step 3.7 Flash

stepfun·stepfun/step-3.7-flash·GA·Open weights·NEW

Newer StepFun flash model for faster agents, coding, and multimodal prompts

At a glance
Official price$0.185 / $1.11 per 1M
Cache read $0.037 · StepFun (Global)
Lowest paid$0.18 / $1.03
Requesty Gateway · 1.1× spread
2 more $0 channels
Context256,000
⚠ Providers report 256,000–262,144; the table below is authoritative
Output limit256,000
Capabilities
ReasoningTool useStructured outputTemperatureAttachments
Modalities
TextImageVideo
Knowledge cutoff2026-03-01
Released / updated2026-05-29 / 2026-05-29

Quality & performanceArtificial Analysis · Intelligence Index v4.1

Intelligence30.9
Value78
Coding39.6
Agentic21.7
Output speed117 tok/s
TTFT2.74 s
Cost per task$0.0913
Value formulaIQ 30.9 ÷ min blended $0.394 = 78

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 16 providers12 with public prices · 2 free · 2 subscription-covered

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NvidiaFirst-partyFree256,00016,384
UnoRouter
step-3.7-flash:free
GatewayFree256,000256,000
RequestyGateway$0.18$1.03$0.036262,144256,000
StepFun (Global)OfficialFirst-party$0.185$1.11$0.037256,000256,000
StepFun (China)OfficialFirst-party$0.185$1.11$0.037256,000256,000
AmbientGateway$0.19$1.14$0.03$0262,144262,144
Deep Infra
stepfun-ai/Step-3.7-Flash
Cloud$0.20$1.15$0.04262,144256,000
Hugging Face
stepfun-ai/Step-3.7-Flash
Cloud$0.20$1.15262,144256,000
OpenRouterGateway$0.20$1.15$0.04262,144256,000
ZenMuxGateway$0.20$1.15256,000256,000

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Requesty$70.47
2StepFun (Global) · Official$74.74
3StepFun (China) · Official$74.74
4Ambient$75.80
5Deep Infra$78.30
6OpenRouter$78.30

2 more channels offer $0 (Nvidia, UnoRouter); free tiers usually have rate limits and no SLA, excluded from ranking.

Switch to Requesty to save $4.27/mo (6%). The gap is small, so staying on the official channel is fine.
Note: this is a gateway; verify availability and rate limits yourself.

Benchmark11 items

NameConditionsScoreMetricSource
SWE-Bench Pro56.3resolve rateSource ↗
SWE-Bench Verified76.5resolvedSource ↗
Terminal-Benchv2.159.6success rateSource ↗
Humanity's Last Examvariant: with tools47.2accuracySource ↗
BrowseComp75.8accuracySource ↗
Toolathlon49.5success rateSource ↗
GDPval45.8wins or tiesSource ↗
ClawEvalv1.167.1pass^3Source ↗
Artificial Analysis Coding Index37.1indexSource ↗
SciCode40percent correctSource ↗
Terminal-Bench Hard35.6success rateSource ↗

The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.

Reasoning control

effort = minimaleffort = loweffort = mediumeffort = higheffort = xhigheffort = maxeffort = nonebudget_tokens

2 / 16 providers expose no reasoning control (reasoning_options: []).

Related models

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$1.11$02026-08-052026-08-212026-08-05 · Input list $0.1852026-08-06 · Input list $0.1852026-08-07 · Input list $0.1852026-08-08 · Input list $0.1852026-08-09 · Input list $0.1852026-08-10 · Input list $0.1852026-08-11 · Input list $0.1852026-08-12 · Input list $0.1852026-08-13 · Input list $0.1852026-08-21 · Input list $0.1852026-08-05 · Output list $1.112026-08-06 · Output list $1.112026-08-07 · Output list $1.112026-08-08 · Output list $1.112026-08-09 · Output list $1.112026-08-10 · Output list $1.112026-08-11 · Output list $1.112026-08-12 · Output list $1.112026-08-13 · Output list $1.112026-08-21 · Output list $1.112026-08-05 · Min blended $0.4162026-08-06 · Min blended $0.4162026-08-07 · Min blended $0.4162026-08-08 · Min blended $0.4162026-08-09 · Min blended $0.4162026-08-10 · Min blended $0.4162026-08-11 · Min blended $0.4162026-08-12 · Min blended $0.4162026-08-13 · Min blended $0.4162026-08-21 · Min blended $0.394