qwen · live price
Qwen2.5 Coder 32B Instruct API pricing
Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). Qwen2.5-Coder brings the following improvements upon CodeQwen1.5: - Significantly improvements in **code generation**, **code reasoning**...
- Input
- $0.66
- per 1M tokens
- Output
- $1
- per 1M tokens
- Cached input
- –
- per 1M tokens
- Context
- 33K
- 29K max output
List price from OpenRouter, which passes qwen's price through without markup. Refreshed hourly. Released November 11, 2024. Model id qwen/qwen-2.5-coder-32b-instruct.
What Qwen2.5 Coder 32B Instruct costs in practice
1,000 chat messages
$1.16
1,000 input + 500 output tokens each
1,000 RAG questions
$5.78
8,000 input + 500 output tokens each
100 coding-agent steps
$2.18
30,000 input + 2,000 output tokens each, no caching
Qwen2.5 Coder 32B Instruct price history
Qwen2.5 Coder 32B Instruct price by host
- Cloudflare$0.66 in · $1 out · 33K ctx · 99.8% uptime 24h
Where to buy qwen credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Qwen2.5 Coder 32B Instruct cost calculator
Cheaper alternatives to Qwen2.5 Coder 32B Instruct
| Model | Input / 1M | Output / 1M |
|---|---|---|
| OpenAI: GPT-6 Luna Pro | $0.1 | $0.5 |
| OpenAI: GPT-6 Luna | $0.1 | $0.5 |
| Z.ai: GLM 5.3 FlashX | $0.37 | $1.25 |
| DeepSeek: DeepSeek V4.1 Flash | $0.03 | $0.5 |
| Z.ai: GLM 5.3 Flash | $0.15 | $0.5 |
| DeepSeek: DeepSeek V4 Flash Vision Exp | $0.216 | $0.647 |
Other qwen models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Qwen2.5 7B Instruct | $0.1 | $0.2 |
| Qwen2.5 72B Instruct | $0.36 | $0.4 |
| Qwen-Plus | $0.26 | $0.78 |
| Qwen2.5 VL 72B Instruct | $0.8 | $1 |
| Qwen3 32B | $0.08 | $0.28 |
| Qwen3 14B | $0.12 | $0.24 |