Inception · live price
Mercury 2.5 API pricing
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
- Input
- $0.04
- per 1M tokens
- Output
- $0.15
- per 1M tokens
- Cached input
- $0.0040
- per 1M tokens
- Context
- 260K
- 66K max output
List price from OpenRouter, which passes Inception's price through without markup. Refreshed hourly. Released September 8, 2026. Model id inception/mercury-2.5.
What Mercury 2.5 costs in practice
1,000 chat messages
$0.12
1,000 input + 500 output tokens each
1,000 RAG questions
$0.40
8,000 input + 500 output tokens each
100 coding-agent steps
$0.15
30,000 input + 2,000 output tokens each, no caching
Mercury 2.5 price history
Mercury 2.5 price by host
- Inception$0.04 in · $0.15 out · 260K ctx · 99.9% uptime 24h
Where to buy Inception credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Mercury 2.5 cost calculator
Cheaper alternatives to Mercury 2.5
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Qwen: Qwen3.7 Flash | $0.03 | $0.13 |
| DeepSeek: DeepSeek V4 Flash 0423 | $0.042 | $0.084 |
| OpenAI: gpt-oss-20b | $0.018 | $0.09 |
| Google: Gemma 3 4B | $0.05 | $0.1 |
| Mistral: Mistral Small 3 | $0.05 | $0.08 |
| Meta: Llama 3.1 8B Instruct | $0.05 | $0.08 |
Other Inception models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Mercury 2 | $0.25 | $0.75 |