GPT-3.5 Turbo
OpenAI · openai/gpt-3.5-turbo
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.
GPT-3.5 Turbo ranks #200 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $1.00 blended per million tokens, it is cheaper than 52% of models with published API pricing. Its 16,385-token context window is 16.0× smaller than the catalog median (262,144 tokens).
texttext->text
Context
Max context: 16385
Max output: 4096
Max output: 4096
Pricing
Input / 1M: 0.50
Output / 1M: 1.50
Blend / 1M: 1.00
Output / 1M: 1.50
Blend / 1M: 1.00
Quality
Quality index: —
Provider
Provider: OpenAI
Moderated: yes
Moderated: yes
Monthly cost at this blended rate
Estimated API spend if every token is billed at $1.00per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $1.00 |
| 10,000,000 | $10.00 |
| 100,000,000 | $100.00 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from OpenAI
- GPT-5.5$17.50 / 1M · Q 81.8
- GPT-5.4$8.75 / 1M · Q 80.6
- GPT-5.2 Chat$7.88 / 1M · Q 76.4
- GPT-5.2$7.88 / 1M · Q 76.4
- GPT-5.1$5.63 / 1M · Q 76.1
Closest API price
- Mistral Large 3 2512$1.00 / 1M
- Morph V3 Fast$1.00 / 1M
- Sonar$1.00 / 1M
- Hermes 3 405B Instruct$1.00 / 1M
- GPT-4.1 Mini$1.00 / 1M