Qwen · live price
Qwen3.5-Flash API pricing
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...
- Input
- $0.065
- per 1M tokens
- Output
- $0.26
- per 1M tokens
- Cached input
- –
- per 1M tokens
- Context
- 1M
- 66K max output
List price from OpenRouter, which passes Qwen's price through without markup. Refreshed hourly. Released February 25, 2026. Model id qwen/qwen3.5-flash-02-23.
What Qwen3.5-Flash costs in practice
1,000 chat messages
$0.20
1,000 input + 500 output tokens each
1,000 RAG questions
$0.65
8,000 input + 500 output tokens each
100 coding-agent steps
$0.25
30,000 input + 2,000 output tokens each, no caching
Qwen3.5-Flash price history
Qwen3.5-Flash price by host
- Alibaba$0.065 in · $0.26 out · 1M ctx · 99.9% uptime 24h
Where to buy Qwen credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Qwen3.5-Flash cost calculator
Cheaper alternatives to Qwen3.5-Flash
| Model | Input / 1M | Output / 1M |
|---|---|---|
| DeepSeek: DeepSeek V4 Flash 0423 | $0.042 | $0.084 |
| Mistral: Ministral 3 3B 2512 | $0.1 | $0.1 |
| OpenAI: gpt-oss-120b | $0.037 | $0.17 |
| OpenAI: gpt-oss-20b | $0.018 | $0.09 |
| Google: Gemma 3 4B | $0.05 | $0.1 |
| Google: Gemma 3 12B | $0.05 | $0.15 |
Other Qwen models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Qwen3.5-122B-A10B | $0.26 | $2.08 |
| Qwen3.5-27B | $0.195 | $1.56 |
| Qwen3.5-35B-A3B | $0.163 | $1.30 |
| Qwen3.5 Plus 2026-02-15 | $0.26 | $1.56 |
| Qwen3.5 397B A17B | $0.55 | $3.50 |
| Qwen3.5-9B | $0.1 | $0.15 |