LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list
mi

ling-3.0-tiny

misc·misc/ling-3.0-tiny·GA·Closed·ling series·NEW
Ling-3.0-tiny is an efficient 7.9B parameter MoE model from inclusionAI with only 1.3B active parameters per token. Built for responsive agents, reliable instruction following and multi turn conversation, with a 256K context window, native function calling, prompt caching and switchable Thinking and Instant modes.

ling-3.0-tiny by misc is currently listed from a single provider. Its reference price is $0 per 1M tokens. It also has 1 free ($0) channel; free tiers usually carry rate limits, and subscription-covered access bills $0 per token only after the subscription fee.

Artificial Analysis rates it 11.1 on the Intelligence Index, with 26.5 for coding and 4.4 for agentic tasks. Median output speed is 57 tokens per second, with 2.52s to the first token.

The context window is 262,144 tokens, with an output limit of 32,768 tokens. It supports reasoning and tool use.

Specs & pricing

Input / output per 1M tokens
Reference price·Requesty
$0 / $0
Blended $0 · Cache read —
Lowest paid
No paid channels
1 more $0 channels
Context
262,144
Output limit
32,768
Knowledge cutoff
—
Released / updated
2026-08-05 / 2026-08-05
Capabilities
✓ Reasoning✓ Tool useStructured output? TemperatureAttachments
Modalities
Text

Quality & performanceArtificial Analysis · Intelligence Index v4.3

Intelligence11.1
Value—
Coding26.5
Agentic4.4
Output speed57 tok/s
TTFT2.52 s

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 1 providers0 with public prices · 1 free

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
RequestyGatewayFree——262,14432,768

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = noneeffort = loweffort = mediumeffort = higheffort = max

Your usage cost

This model is $0 across all listed channels (Requesty).

Price historyone sample accumulated per data sync

Input list $0Output list $0Min blended —

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-19). Each future sync adds a point, and once accumulated a line is drawn here.

Artificial Analysis evaluations6 items

GPQA Diamond73.4%
Humanity's Last Exam9.3%
SciCode24.2%
Terminal-Bench 2.127.7%
τ³-Bench Banking20.8%
AA-LCR60.3%

Individual evaluations run by Artificial Analysis, on the same reasoning tier as the intelligence score above. Each benchmark has its own task set and harness, so rows are not comparable with one another. The Intelligence Index above draws on a different, newer set of evaluations.

Related models

inLing 3.1 Flashsame series$0.075 / $0.22inLing 3.1 Flash (Free)same series$0 / $0inLing 3.0 Flash VLsame series$0.075 / $0.22
Data partly from models.dev (MIT) · Requesty official docs ↗