ling-3.0-tiny
Ling-3.0-tiny is an efficient 7.9B parameter MoE model from inclusionAI with only 1.3B active parameters per token. Built for responsive agents, reliable instruction following and multi turn conversation, with a 256K context window, native function calling, prompt caching and switchable Thinking and Instant modes.
Quality & performanceArtificial Analysis · Intelligence Index v4.1
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 1 providers0 with public prices · 1 free
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Requesty | Gateway | Free | — | — | 262,144 | 32,768 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
This model is $0 across all listed channels (Requesty).
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
Related models
Price historyone sample accumulated per data sync
Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-19). Each future sync adds a point, and once accumulated a line is drawn here.