openai

GPT-3.5 Turbo 16k pricing & benchmarks

Released Aug 28, 2023 · served by 2 providers

Input / 1M tokens$3.00
Output / 1M tokens$4.00
Context window16K
Cached input
1M in + 300K out$4.20

What GPT-3.5 Turbo 16k is

This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up...

Providers and prices for GPT-3.5 Turbo 16k

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
OpenAIcheapest $3.00 $4.00 16K 100.0%
Azure $3.00 $4.00 16K 99.6%

FAQ

How much does the GPT-3.5 Turbo 16k API cost?

$3.00 per 1M input tokens and $4.00 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $4.20.

Is GPT-3.5 Turbo 16k free?

No. It is a paid model starting at $3.00 per 1M input tokens, though some providers offer trial credits.

Which provider is cheapest for GPT-3.5 Turbo 16k?

OpenAI at $3.00 input / $4.00 output per 1M tokens.

Can I self-host GPT-3.5 Turbo 16k?

No. GPT-3.5 Turbo 16k is closed-weight and only available through hosted APIs.