GPT-3.5 Turbo (older v0613) pricing & benchmarks
Released Jan 25, 2024 · served by 1 provider
What GPT-3.5 Turbo (older v0613) is
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks.
- Model ID:
openai/gpt-3.5-turbo-0613 - Modality: text->text
- Knowledge cutoff: 2021-09-30
- Tool calling: supported · Structured output: supported
Providers and prices for GPT-3.5 Turbo (older v0613)
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| Azurecheapest | $1.00 | $2.00 | 4K | — | 100.0% | — |
FAQ
How much does the GPT-3.5 Turbo (older v0613) API cost?
$1.00 per 1M input tokens and $2.00 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $1.60.
Is GPT-3.5 Turbo (older v0613) free?
No. It is a paid model starting at $1.00 per 1M input tokens, though some providers offer trial credits.
Which provider is cheapest for GPT-3.5 Turbo (older v0613)?
Azure at $1.00 input / $2.00 output per 1M tokens.
Can I self-host GPT-3.5 Turbo (older v0613)?
No. GPT-3.5 Turbo (older v0613) is closed-weight and only available through hosted APIs.