← Model list
Inkling Small Thinking
thinkingmachines·thinkingmachines/inkling-small-thinking·GA·Open weights·ling series·Video gen·NEW
Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio
At a glance
Reference price$0.50 / $1.20 per 1M
Cache read $0.10 · NanoGPT
Lowest paid$0.50 / $1.20
NanoGPT Gateway
Context524,288
Output limit32,768
Capabilities
✓ Reasoning✓ Tool useStructured output✓ Temperature✓ Attachments
Modalities
TextImage
Knowledge cutoff—
Released / updated2026-07-30 / 2026-07-30
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 1 providers1 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NanoGPT thinkingmachines/Inkling-Small:thinking | Gateway | $0.50 | $1.20 | $0.10 | — | 524,288 | 32,768 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1NanoGPT$112.00
The cheapest paid channel is the only channel.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
effort = minimaleffort = loweffort = mediumeffort = higheffort = max
Related models
Ling 3.0 Flashsame series$0.075 / $0.22ling-3.0-tinysame series$0 / $0InclusionAI Ling 3.0 Flash (DeepInfra)same series$0.06 / $0.18DeepSeek V4 Flashcheaper alternative$0.14 / $0.28GPT-5 Nanocheaper alternative$0.05 / $0.40MiMo-V2.5cheaper alternative$0.14 / $0.28Gemini 2.5 Flash-Litecheaper alternative$0.10 / $0.40
Price historyone sample accumulated per data sync
Input listOutput listMin blended