NVIDIA Nemotron Nano v2 12B is a 12-billion-parameter multimodal reasoning model designed for advanced video understanding, document intelligence, and visual reasoning, built with a hybrid Transformer-Mamba architecture for high efficiency and low latency.
nemotron-nano-v2-12b by misc is currently listed from a single provider. Its reference price is $0.24 per 1M input tokens and $0.707 per 1M output tokens.
The context window is 128,000 tokens, with an output limit of 128,000 tokens. It supports reasoning, tool use, and structured output. Accepted input modalities are Text and Image.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Cortecs | Gateway | $0.24 | $0.71 | — | — | 128,000 | 128,000 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
1 / 1 providers expose no reasoning control (reasoning_options: []).