inclusionAI · live price
Ling 3.0 Flash VL API pricing
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
- Input
- $0.021
- per 1M tokens
- Output
- $0.062
- per 1M tokens
- Cached input
- $0.0042
- per 1M tokens
- Context
- 262K
- 33K max output
List price from OpenRouter, which passes inclusionAI's price through without markup. Refreshed hourly. Released September 10, 2026. Model id inclusionai/ling-3.0-flash-vl.
What Ling 3.0 Flash VL costs in practice
1,000 chat messages
$0.05
1,000 input + 500 output tokens each
1,000 RAG questions
$0.20
8,000 input + 500 output tokens each
100 coding-agent steps
$0.08
30,000 input + 2,000 output tokens each, no caching
Ling 3.0 Flash VL price history
Ling 3.0 Flash VL price by host
- Novita$0.021 in · $0.062 out · 262K ctx · 99.8% uptime 24h
- DeepInfra$0.06 in · $0.18 out · 131K ctx · 99.9% uptime 24h
Where to buy inclusionAI credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Ling 3.0 Flash VL cost calculator
Cheaper alternatives to Ling 3.0 Flash VL
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Mistral: Mistral Nemo | $0.019 | $0.03 |
Other inclusionAI models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Ling 3.0 Flash Fin | $0.06 | $0.18 |
| Ling 3.0 Flash | $0.021 | $0.063 |