Qwen · live price
Qwen3 VL 8B Thinking API pricing
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...
- Input
- $0.18
- per 1M tokens
- Output
- $2.10
- per 1M tokens
- Cached input
- –
- per 1M tokens
- Context
- 131K
- 33K max output
List price from OpenRouter, which passes Qwen's price through without markup. Refreshed hourly. Released October 14, 2025. Model id qwen/qwen3-vl-8b-thinking.
What Qwen3 VL 8B Thinking costs in practice
1,000 chat messages
$1.23
1,000 input + 500 output tokens each
1,000 RAG questions
$2.49
8,000 input + 500 output tokens each
100 coding-agent steps
$0.96
30,000 input + 2,000 output tokens each, no caching
Qwen3 VL 8B Thinking price history
Qwen3 VL 8B Thinking price by host
- Alibaba$0.18 in · $2.10 out · 131K ctx · 99.9% uptime 24h
Where to buy Qwen credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Qwen3 VL 8B Thinking cost calculator
Cheaper alternatives to Qwen3 VL 8B Thinking
| Model | Input / 1M | Output / 1M |
|---|---|---|
| OpenAI: GPT-6 Luna Pro | $0.1 | $0.5 |
| OpenAI: GPT-6 Luna | $0.1 | $0.5 |
| Z.ai: GLM 5.3 FlashX | $0.37 | $1.25 |
| DeepSeek: DeepSeek V4.1 Flash | $0.03 | $0.5 |
| Z.ai: GLM 5.3 Flash | $0.15 | $0.5 |
| DeepSeek: DeepSeek V4 Flash Vision Exp | $0.216 | $0.647 |
Other Qwen models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Qwen3 VL 30B A3B Instruct | $0.15 | $0.6 |
| Qwen3 VL 235B A22B Instruct | $0.21 | $1.90 |
| Qwen3 Coder Flash | $0.195 | $0.975 |
| Qwen3 Next 80B A3B Thinking | $0.15 | $1.20 |
| Qwen3 Next 80B A3B Instruct | $0.1 | $1.10 |
| Qwen3 Coder 30B A3B Instruct | $0.07 | $0.28 |