← Model list
inclusionAI/Ling-flash-2.0
inclusionai·inclusionai/ling-flash-2.0·GA·Closed·ling series
Efficient model for low-latency assistance, extraction, and routine automation
At a glance
Reference price$0.14 / $0.57 per 1M
Cache read — · SiliconFlow (China)
Lowest paid$0.14 / $0.57
SiliconFlow Cloud · 1× spread
Context131,000
Output limit131,000
Capabilities
Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
Modalities
Text
Knowledge cutoff—
Released / updated2025-09-18 / 2025-11-25
Quality & performanceArtificial Analysis · Intelligence Index v4.1
Intelligence9.6
Value39
Coding—
Agentic—
Output speed84 tok/s
TTFT2.25 s
Value formulaIQ 9.6 ÷ min blended $0.247 = 39
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 2 providers2 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| SiliconFlow inclusionAI/Ling-flash-2.0 | Cloud | $0.14 | $0.57 | — | — | 131,000 | 131,000 | |
| SiliconFlow (China) inclusionAI/Ling-flash-2.0 | Cloud | $0.14 | $0.57 | — | — | 131,000 | 131,000 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1SiliconFlow$56.50
2SiliconFlow (China)$56.50
The cheapest paid channel is the only channel.
Related models
Ling 3.0 Flashsame series$0.075 / $0.22ling-3.0-tinysame series$0 / $0InclusionAI Ling 3.0 Flash (DeepInfra)same series$0.06 / $0.18Llama 3.2 1B Instructcheaper alternative$0.10 / $0.201Llama 4 Maverick 17B 128E Instruct FP8cheaper alternative$0 / $0Mistral Nemo Instruct 2407cheaper alternative$0.145 / $0.145Phi 4 Multimodalcheaper alternative$0.08 / $0.32
Price historyone sample accumulated per data sync
Input listOutput listMin blended