Dense 2B-class open-source model for on-device and resource-constrained use, with native long-context support, tool calling, and agentic tasks
MiniCPM5-2B by openbmb is currently listed from a single provider. Its reference price is $0.124 per 1M input tokens and $0.743 per 1M output tokens.
Artificial Analysis rates it 12.5 on the Intelligence Index, with 14.5 for coding and 6.7 for agentic tasks. Against its lowest blended price of $0.279 per 1M, that is roughly 45 index points per dollar, which is the value ratio the leaderboards rank on.
The context window is 131,072 tokens, with an output limit of 131,072 tokens. It supports reasoning and tool use. The weights are open, so it can also be self-hosted or served through a gateway of your choice.
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| AMD MiniCPM5-2B | Cloud | $0.124 | $0.743 | $0.124 | — | 131,072 | 131,072 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
1 / 1 providers expose no reasoning control (reasoning_options: []).
Price history accumulates from each data sync; currently only 1 sample(s) (2026-09-15). Each future sync adds a point, and once accumulated a line is drawn here.
Individual evaluations run by Artificial Analysis, on the same reasoning tier as the intelligence score above. Each benchmark has its own task set and harness, so rows are not comparable with one another. The Intelligence Index above draws on a different, newer set of evaluations.