Pareto Inference is a third-party gateway that routes to many upstream models through one API. Gateway quotes are often among the lowest on the site, but availability, rate limits and routing can vary between models, so a low headline price is worth verifying against your own workload.
It lists 1 model on this page. Coverage spans text and chat (1). The lowest-priced model hosted here is GLM-5.3-Flash.
Pareto Inference is the lowest paid channel for 1 model that has at least two competing paid hosts, which is where its pricing genuinely leads rather than being the only option.
| Cheapest here? | |||||
|---|---|---|---|---|---|
| GLM-5.3-Flashz-ai/glm-5.3-flash | $0.03 | $0.10 | $0.006 | 1,000,000 | Lowest anywhere |