inclusionAI · live price
Ling 3.0 Flash Fin API pricing
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
- Input
- $0.06
- per 1M tokens
- Output
- $0.18
- per 1M tokens
- Cached input
- $0.012
- per 1M tokens
- Context
- 262K
- 236K max output
List price from OpenRouter, which passes inclusionAI's price through without markup. Refreshed hourly. Released August 27, 2026. Model id inclusionai/ling-3.0-flash-fin.
What Ling 3.0 Flash Fin costs in practice
1,000 chat messages
$0.15
1,000 input + 500 output tokens each
1,000 RAG questions
$0.57
8,000 input + 500 output tokens each
100 coding-agent steps
$0.22
30,000 input + 2,000 output tokens each, no caching
Ling 3.0 Flash Fin price history
Ling 3.0 Flash Fin price by host
- Novita$0.042 in · $0.123 out · 262K ctx · 100.0% uptime 24h
- DeepInfra$0.06 in · $0.18 out · 262K ctx · 99.9% uptime 24h
Where to buy inclusionAI credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Ling 3.0 Flash Fin cost calculator
Cheaper alternatives to Ling 3.0 Flash Fin
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Qwen: Qwen3.7 Flash | $0.03 | $0.13 |
| DeepSeek: DeepSeek V4 Flash 0423 | $0.042 | $0.084 |
| OpenAI: gpt-oss-120b | $0.037 | $0.17 |
| OpenAI: gpt-oss-20b | $0.018 | $0.09 |
| Qwen: Qwen3 30B A3B Instruct 2507 | $0.048 | $0.193 |
| Google: Gemma 3 4B | $0.05 | $0.1 |
Other inclusionAI models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Ling 3.0 Flash VL | $0.021 | $0.062 |
| Ling 3.0 Flash | $0.021 | $0.063 |