LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list
mi

ling-3.0-tiny

misc·misc/ling-3.0-tiny·GA·Closed·ling series·NEW
Ling-3.0-tiny is an efficient 7.9B parameter MoE model from inclusionAI with only 1.3B active parameters per token. Built for responsive agents, reliable instruction following and multi turn conversation, with a 256K context window, native function calling, prompt caching and switchable Thinking and Instant modes.

ling-3.0-tiny by misc is currently listed from a single provider. Its reference price is $0 per 1M tokens. It also has 1 free ($0) channel; free tiers usually carry rate limits, and subscription-covered access bills $0 per token only after the subscription fee.

Artificial Analysis rates it 24.5 on the Intelligence Index, with 26.5 for coding and 16 for agentic tasks. Median output speed is 161 tokens per second, with 2.6s to the first token.

The context window is 262,144 tokens, with an output limit of 32,768 tokens. It supports reasoning and tool use.

Specs & pricing

Input / output per 1M tokens
Reference price·Requesty
$0 / $0
Blended $0 · Cache read —
Lowest paid
No paid channels
1 more $0 channels
Context
262,144
Output limit
32,768
Knowledge cutoff
—
Released / updated
2026-08-05 / 2026-08-05
Capabilities
✓ Reasoning✓ Tool useStructured output? TemperatureAttachments
Modalities
Text

Quality & performanceArtificial Analysis · Intelligence Index v4.1

Intelligence24.5
Value—
Coding26.5
Agentic16
Output speed161 tok/s
TTFT2.6 s
Cost per task$0

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 1 providers0 with public prices · 1 free

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
RequestyGatewayFree——262,14432,768

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

This model is $0 across all listed channels (Requesty).

Reasoning control

effort = noneeffort = loweffort = mediumeffort = higheffort = maxbudget_tokens

Related models

inLing 3.0 Flash Sante (Free)same series$0 / $0inLing 3.0 Flash Santesame series$0 / $0inLing 3.0 Flash Finsame series$0.06 / $0.18

Price historyone sample accumulated per data sync

Input list $0Output list $0Min blended —

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-19). Each future sync adds a point, and once accumulated a line is drawn here.

Data partly from models.dev (MIT) · Requesty official docs ↗