GPT-3.5 Turbo pricing & benchmarks
Released May 28, 2023 · served by 1 provider
What GPT-3.5 Turbo is
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks.
- Model ID:
openai/gpt-3.5-turbo - Modality: text->text
- Knowledge cutoff: 2021-09-30
- Tool calling: supported · Structured output: supported
Providers and prices for GPT-3.5 Turbo
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| OpenAIcheapest | $0.500 | $1.50 | 16K | — | 100.0% | — |
FAQ
How much does the GPT-3.5 Turbo API cost?
$0.500 per 1M input tokens and $1.50 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $0.95.
Is GPT-3.5 Turbo free?
No. It is a paid model starting at $0.500 per 1M input tokens, though some providers offer trial credits.
Which provider is cheapest for GPT-3.5 Turbo?
OpenAI at $0.500 input / $1.50 output per 1M tokens.
Can I self-host GPT-3.5 Turbo?
No. GPT-3.5 Turbo is closed-weight and only available through hosted APIs.