Open-weight GPT model for self-hosted reasoning and instruction-following workloads
GPT OSS 120B High Throughput by openai is currently listed from a single provider. Its reference price is $0.09 per 1M input tokens and $0.36 per 1M output tokens.
The context window is 131,072 tokens, with an output limit of 16,384 tokens. It supports reasoning and tool use. The weights are open, so it can also be self-hosted or served through a gateway of your choice.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Clarifai | Cloud | $0.09 | $0.36 | — | — | 131,072 | 16,384 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.