Nemotron middle tier for collaborative agents and high-volume reasoning workloads
nemotron-3-super by nvidia is offered by 2 providers on this page. Public prices are shown for 1 of them. Its official list price is $0.20 per 1M input tokens and $0.80 per 1M output tokens.
The context window is 262,144 tokens, with an output limit of 262,144 tokens. It supports reasoning and tool use. The weights are open, so it can also be self-hosted or served through a gateway of your choice. Its training knowledge cuts off at 2024-04.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NvidiaOfficial nvidia/nemotron-3-super-120b-a12b | First-party | $0.20 | $0.80 | — | — | 262,144 | 262,144 | |
| Ollama Cloud | Cloud | — | — | — | — | 262,144 | 65,536 | hostTable.undisclosed |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Price history accumulates from each data sync; currently only 1 sample(s) (2026-09-04). Each future sync adds a point, and once accumulated a line is drawn here.