Ling-3.0-flash Thinking enables visible reasoning on inclusionAI's token-efficient 124B-parameter Mixture-of-Experts model for harder coding, tool use, planning, and production-scale agent workflows.
Ling 3.0 Flash Thinking by inclusionai is currently listed from a single provider. Its reference price is $0.075 per 1M input tokens and $0.22 per 1M output tokens.
The context window is 262,144 tokens, with an output limit of 32,768 tokens. It supports reasoning and tool use.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NanoGPT inclusionai/ling-3.0-flash:thinking | Gateway | $0.075 | $0.22 | $0.015 | — | 262,144 | 32,768 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.