GPT-3.5 Turbo pricing
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021. This page tracks 6 listings in total. Highlighted lows are $0.500 per million input and $1.50 per million output (see table for which seller matches each).
Pricing across providers
Every row is a seller of GPT-3.5 Turbo with token pricing we track. The cheapest input in this snapshot is from Azure. The bar chart shows the same input and output dollars per million for a quick scan.
| Provider | Input / 1M | Output / 1M | Cached input | Batch |
|---|---|---|---|---|
A Azure | $0.500 | $1.50 | — | — |
GC Github Copilot | N/A | N/A | — | — |
O Openrouter | $0.500 | $1.50 | — | — |
TC Text Completion Openai | $1.50 | $2.00 | — | — |
VA Vercel Ai Gateway | $0.500 | $1.50 | — | — |
O OpenAInative | $0.500 | $1.50 | — | — |
Input vs output · per provider
Cost calculator
The calculator uses the same dollars per million tokens as the table. Adjust sliders to see how GPT-3.5 Turbo cost scales with traffic.
Provider
0.050000¢ / req
0.075000¢ / req
Model specifications
Quick spec sheet for GPT-3.5 Turbo before you dive back into pricing. Reported under OpenAI.
- Context window
- 16,385 tokens
- Max output
- 4,096 tokens
- Vision (images)
- No
- Tool / function calling
- Yes
- Streaming
- No
- Released
- May 2023
- Primary provider
- OpenAI
- Model family
- N/A
Compare GPT-3.5 Turbo
Open a pair page to see GPT-3.5 Turbo next to another model with a shared provider matrix. 6 shortcuts below.
Locked
Compare with
Pick a model on both sides.
Popular GPT-3.5 Turbo comparisons
- GPT-3.5 Turbo vs Claude Sonnet 4.6
GPT-3.5 Turbo 90% cheaper on output
- GPT-3.5 Turbo vs Gemini 2.0 Flash
Gemini 2.0 Flash 73% cheaper on output
- GPT-3.5 Turbo vs Llama 3.1 70B
Compare pricing side by side
- GPT-3.5 Turbo vs Mistral 7B
Mistral 7B 83% cheaper on output
- GPT-3.5 Turbo vs GLM 4.7
GPT-3.5 Turbo 31% cheaper on output
- GPT-3.5 Turbo vs Amazon Nova Lite V1 0
Compare pricing side by side
Frequently asked questions
Read these after the table if you want plain language around GPT-3.5 Turbo rates. The short model note from our index: GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.
Related models
Models close to GPT-3.5 Turbo by family, provider, and context window — each with live pricing in our catalog.