openai

GPT-3.5 Turbo (older v0613) pricing & benchmarks

Released Jan 25, 2024 · served by 1 provider

Input / 1M tokens$1.00
Output / 1M tokens$2.00
Context window4K
Cached input
1M in + 300K out$1.60

What GPT-3.5 Turbo (older v0613) is

GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks.

Providers and prices for GPT-3.5 Turbo (older v0613)

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
Azurecheapest $1.00 $2.00 4K 100.0%

FAQ

How much does the GPT-3.5 Turbo (older v0613) API cost?

$1.00 per 1M input tokens and $2.00 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $1.60.

Is GPT-3.5 Turbo (older v0613) free?

No. It is a paid model starting at $1.00 per 1M input tokens, though some providers offer trial credits.

Which provider is cheapest for GPT-3.5 Turbo (older v0613)?

Azure at $1.00 input / $2.00 output per 1M tokens.

Can I self-host GPT-3.5 Turbo (older v0613)?

No. GPT-3.5 Turbo (older v0613) is closed-weight and only available through hosted APIs.