Mellum2-12B-A2.5B-Instruct is a fast MoE model with 131K context built for coding, tool use, and low-latency AI workflows.
Mellum2 12B A2.5B by jetbrains is currently listed from a single provider. Its reference price is $0.05 per 1M input tokens and $0.10 per 1M output tokens.
The context window is 131,072 tokens, with an output limit of 131,072 tokens. It supports tool use and structured output. The weights are open, so it can also be self-hosted or served through a gateway of your choice.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Weights & Biases JetBrains/Mellum2-12B-A2.5B-Instruct | Cloud | $0.05 | $0.10 | $0.05 | — | 131,072 | 131,072 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.