Qwen · live price
Qwen3 32B API pricing
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
- Input
- $0.08
- per 1M tokens
- Output
- $0.28
- per 1M tokens
- Cached input
- –
- per 1M tokens
- Context
- 131K
- 16K max output
List price from OpenRouter, which passes Qwen's price through without markup. Refreshed hourly. Released April 28, 2025. Model id qwen/qwen3-32b.
What Qwen3 32B costs in practice
1,000 chat messages
$0.22
1,000 input + 500 output tokens each
1,000 RAG questions
$0.78
8,000 input + 500 output tokens each
100 coding-agent steps
$0.30
30,000 input + 2,000 output tokens each, no caching
Qwen3 32B price history
Qwen3 32B price by host
- DeepInfra$0.08 in · $0.28 out · 41K ctx · 100.0% uptime 24h
- SiliconFlow$0.14 in · $0.57 out · 131K ctx · 99.9% uptime 24h
Where to buy Qwen credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Qwen3 32B cost calculator
Cheaper alternatives to Qwen3 32B
| Model | Input / 1M | Output / 1M |
|---|---|---|
| DeepSeek: DeepSeek V4 Flash 0423 | $0.042 | $0.084 |
| Mistral: Ministral 3 3B 2512 | $0.1 | $0.1 |
| OpenAI: gpt-oss-120b | $0.037 | $0.17 |
| OpenAI: gpt-oss-20b | $0.018 | $0.09 |
| Google: Gemma 3 4B | $0.05 | $0.1 |
| Google: Gemma 3 12B | $0.05 | $0.15 |
Other Qwen models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Qwen3 14B | $0.12 | $0.24 |
| Qwen3 30B A3B | $0.12 | $0.5 |
| Qwen3 235B A22B Instruct 2507 | $0.087 | $0.35 |
| Qwen3 Coder 480B A35B | $0.3 | $1 |
| Qwen2.5 VL 72B Instruct | $0.8 | $1 |
| Qwen-Plus | $0.26 | $0.78 |