Ling-3.0-tiny is an efficient 7.9B parameter MoE model from inclusionAI with only 1.3B active parameters per token. Built for responsive agents, reliable instruction following and multi turn conversation, with a 256K context window, native function calling, prompt caching and switchable Thinking and Instant modes.
ling-3.0-tiny by misc is currently listed from a single provider. Its reference price is $0 per 1M tokens. It also has 1 free ($0) channel; free tiers usually carry rate limits, and subscription-covered access bills $0 per token only after the subscription fee.
Artificial Analysis rates it 11.1 on the Intelligence Index, with 26.5 for coding and 4.4 for agentic tasks. Median output speed is 57 tokens per second, with 2.52s to the first token.
The context window is 262,144 tokens, with an output limit of 32,768 tokens. It supports reasoning and tool use.
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Requesty | Gateway | Free | — | — | 262,144 | 32,768 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
This model is $0 across all listed channels (Requesty).
Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-19). Each future sync adds a point, and once accumulated a line is drawn here.
Individual evaluations run by Artificial Analysis, on the same reasoning tier as the intelligence score above. Each benchmark has its own task set and harness, so rows are not comparable with one another. The Intelligence Index above draws on a different, newer set of evaluations.