← Model list
Ling-2.6-flash
inclusionai·inclusionai/ling-2.6-flash·GA·Open weights·ling series
Efficient model for low-latency assistance, extraction, and routine automation
At a glance
Reference price$0.10 / $0.30 per 1M
Cache read $0.02 · NovitaAI
Lowest paid$0.01 / $0.03
OpenRouter Gateway · 10× spread
Context262,144
Output limit32,768
Capabilities
Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
Text
Knowledge cutoff—
Released / updated2026-04-21 / 2026-04-24
Quality & performanceArtificial Analysis · Intelligence Index v4.1
Intelligence14.2
Value947
Coding25.3
Agentic2.3
Output speed91 tok/s
TTFT1.12 s
Value formulaIQ 14.2 ÷ min blended $0.015 = 947
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 4 providers4 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| OpenRouter | Gateway | $0.01 | $0.03 | $0.002 | — | 262,144 | 32,768 | |
| Requesty | Gateway | $0.09 | $0.27 | — | — | 262,144 | 262,144 | |
| NanoGPT | Gateway | $0.10 | $0.30 | $0.02 | — | 262,144 | 32,768 | |
| NovitaAI | Gateway | $0.10 | $0.30 | $0.02 | — | 262,144 | 32,768 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1OpenRouter$2.54
2NanoGPT$25.40
3NovitaAI$25.40
4Requesty$31.50
The cheapest paid channel is the only channel.
Related models
Ling 3.0 Flashsame series$0.075 / $0.22ling-3.0-tinysame series$0 / $0InclusionAI Ling 3.0 Flash (DeepInfra)same series$0.06 / $0.18Lyria 3 Clip Previewcheaper alternative$0 / $0Granite 4.1 8Bcheaper alternative$0.05 / $0.10Lyria 3 Pro Previewcheaper alternative$0 / $0Google Gemma 3 12Bcheaper alternative$0.05 / $0.15
Price historyone sample accumulated per data sync
Input listOutput listMin blended