OpenAI · live price
GPT-3.5 Turbo 16k API pricing
This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up...
- Input
- $3
- per 1M tokens
- Output
- $4
- per 1M tokens
- Cached input
- –
- per 1M tokens
- Context
- 16K
- 4K max output
List price from OpenRouter, which passes OpenAI's price through without markup. Refreshed hourly. Released August 28, 2023. Model id openai/gpt-3.5-turbo-16k.
What GPT-3.5 Turbo 16k costs in practice
1,000 chat messages
$5.00
1,000 input + 500 output tokens each
1,000 RAG questions
$26.00
8,000 input + 500 output tokens each
100 coding-agent steps
$9.80
30,000 input + 2,000 output tokens each, no caching
GPT-3.5 Turbo 16k price history
GPT-3.5 Turbo 16k price by host
- Azure$3 in · $4 out · 16K ctx · 99.9% uptime 24h
- OpenAI$3 in · $4 out · 16K ctx · 99.8% uptime 24h
Where to buy OpenAI credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
GPT-3.5 Turbo 16k cost calculator
Cheaper alternatives to GPT-3.5 Turbo 16k
| Model | Input / 1M | Output / 1M |
|---|---|---|
| SpaceXAI: Grok 4.7 | $2 | $6 |
| Qwen: Qwen3.8 Omni Flash | $0.15 | $0.47 |
| Z.ai: GLM 5.3 FlashX | $0.37 | $1.25 |
| DeepSeek: DeepSeek V4.1 Flash | $0.03 | $0.5 |
| Qwen: Qwen3.8 Max (0902) | $2 | $6 |
| Google: Gemini 3.8 Flash | $0.75 | $3.75 |
Other OpenAI models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| GPT-3.5 Turbo Instruct | $1.50 | $2 |
| GPT-3.5 Turbo | $0.5 | $1.50 |
| GPT-4 | $30 | $60 |
| GPT-3.5 Turbo (older v0613) | $1 | $2 |
| GPT-4 Turbo | $10 | $30 |
| GPT-4o | $2.50 | $10 |