Ling-3.0-tiny is an efficient 7.9B parameter MoE model from inclusionAI with only 1.3B active parameters per token. Built for responsive agents, reliable instruction following and multi turn conversation, with a 256K context window, native function calling, prompt caching and switchable Thinking and Instant modes.
ling-3.0-tiny by misc is currently listed from a single provider. Its reference price is $0 per 1M tokens. It also has 1 free ($0) channel; free tiers usually carry rate limits, and subscription-covered access bills $0 per token only after the subscription fee.
Artificial Analysis rates it 24.5 on the Intelligence Index, with 26.5 for coding and 16 for agentic tasks. Median output speed is 161 tokens per second, with 2.6s to the first token.
The context window is 262,144 tokens, with an output limit of 32,768 tokens. It supports reasoning and tool use.
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Requesty | Gateway | Free | — | — | 262,144 | 32,768 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
This model is $0 across all listed channels (Requesty).
Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-19). Each future sync adds a point, and once accumulated a line is drawn here.