GPT-3.5 Turbo 16k pricing & benchmarks
Released Aug 28, 2023 · served by 2 providers
What GPT-3.5 Turbo 16k is
This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up...
- Model ID:
openai/gpt-3.5-turbo-16k - Modality: text->text
- Knowledge cutoff: 2021-09-30
- Tool calling: supported · Structured output: supported
Providers and prices for GPT-3.5 Turbo 16k
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| OpenAIcheapest | $3.00 | $4.00 | 16K | — | 100.0% | — |
| Azure | $3.00 | $4.00 | 16K | — | 99.6% | — |
FAQ
How much does the GPT-3.5 Turbo 16k API cost?
$3.00 per 1M input tokens and $4.00 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $4.20.
Is GPT-3.5 Turbo 16k free?
No. It is a paid model starting at $3.00 per 1M input tokens, though some providers offer trial credits.
Which provider is cheapest for GPT-3.5 Turbo 16k?
OpenAI at $3.00 input / $4.00 output per 1M tokens.
Can I self-host GPT-3.5 Turbo 16k?
No. GPT-3.5 Turbo 16k is closed-weight and only available through hosted APIs.