← Model list
Inkling Thinking
thinkingmachines·thinkingmachines/inkling-thinking·GA·Open weights·ling series·Video gen·NEW
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
At a glance
Reference price$1.00 / $4.05 per 1M
Cache read $0.17 · NanoGPT
Lowest paid$1.00 / $4.05
NanoGPT Gateway
Context1,048,000
Output limit32,768
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImageAudio
Knowledge cutoff—
Released / updated2026-07-15 / 2026-07-15
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 1 providers1 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NanoGPT thinkingmachines/inkling:thinking | Gateway | $1.00 | $4.05 | $0.17 | — | 1,048,000 | 32,768 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1NanoGPT$302.90
The cheapest paid channel is the only channel.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
effort = minimaleffort = loweffort = mediumeffort = higheffort = xhigh
Related models
Ling 3.0 Flashsame series$0.075 / $0.22ling-3.0-tinysame series$0 / $0InclusionAI Ling 3.0 Flash (DeepInfra)same series$0.06 / $0.18DeepSeek V4 Procheaper alternative$0.435 / $0.87DeepSeek V4 Flashcheaper alternative$0.14 / $0.28DeepSeek V4 Flash 0731cheaper alternative$0.479 / $1.44MiniMax-M3cheaper alternative$0.30 / $1.20
Price historyone sample accumulated per data sync
Input listOutput listMin blended