IFM's open-weight 7B model for coding, reasoning, and chat. Heabsy serves an FP8 deployment on a dedicated RTX 5090 with speculative decoding in the United States, with a 131K context window and an 8K output limit. Zero data retention is not verified for this deployment.
K2-Horizon-7B by misc is currently listed from a single provider. Its reference price is $0.05 per 1M input tokens and $0.20 per 1M output tokens.
The context window is 131,072 tokens, with an output limit of 8,192 tokens. It supports reasoning and tool use. The weights are open, so it can also be self-hosted or served through a gateway of your choice.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NanoGPT | Gateway | $0.05 | $0.20 | $0.04 | — | 131,072 | 8,192 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Price history accumulates from each data sync; currently only 1 sample(s) (2026-09-10). Each future sync adds a point, and once accumulated a line is drawn here.