← Model list
GPT-3.5 Turbo (Azure)
openai·openai/gpt-3.5-turbo-3·GA·Closed·gpt series
Compact GPT model for low-latency assistance and high-volume workloads
At a glance
Reference price$0.50 / $1.50 per 1M
Cache read — · LLM Gateway
Lowest paid$0.50 / $1.50
LLM Gateway Gateway
Context16,385
Output limit4,096
Capabilities
Reasoning✓ Tool useStructured output✓ TemperatureAttachments
Modalities
Text
Knowledge cutoff2021-09-01
Released / updated2023-03-01 / 2023-11-06
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 1 providers1 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| LLM Gateway azure/gpt-3.5-turbo | Gateway | $0.50 | $1.50 | — | — | 16,385 | 4,096 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1LLM Gateway$175.00
The cheapest paid channel is the only channel.
Related models
RouteLLMsame series$3.00 / $15.00GPT-Realtime-2.1same series$4.00 / $24.00NanoGPT Helpsame series$0 / $0Qwen3-Next 80B-A3B Instructcheaper alternative$0.144 / $0.574Qwen3-Coder 30B-A3B Instructcheaper alternative$0.216 / $0.861Llama-3.1-8B-Instructcheaper alternative$0.15 / $0.45Qwen3 Coder Flashcheaper alternative$0.144 / $0.574
Price historyone sample accumulated per data sync
Input list $0.50Output list $1.50Min blended $0.75
Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-21). Each future sync adds a point, and once accumulated a line is drawn here.