IBM · live price
Granite 4.2 8B API pricing
Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...
- Input
- $0.06
- per 1M tokens
- Output
- $0.25
- per 1M tokens
- Cached input
- $0.015
- per 1M tokens
- Context
- 131K
- 118K max output
List price from OpenRouter, which passes IBM's price through without markup. Refreshed hourly. Released August 31, 2026. Model id ibm-granite/granite-4.2-8b.
What Granite 4.2 8B costs in practice
1,000 chat messages
$0.19
1,000 input + 500 output tokens each
1,000 RAG questions
$0.61
8,000 input + 500 output tokens each
100 coding-agent steps
$0.23
30,000 input + 2,000 output tokens each, no caching
Granite 4.2 8B price history
Granite 4.2 8B price by host
- CoreWeave$0.1 in · $0.15 out · 131K ctx · 100.0% uptime 24h
- DeepInfra$0.06 in · $0.25 out · 131K ctx · 100.0% uptime 24h
Where to buy IBM credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Granite 4.2 8B cost calculator
Cheaper alternatives to Granite 4.2 8B
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Qwen: Qwen3.7 Flash | $0.03 | $0.13 |
| DeepSeek: DeepSeek V4 Flash 0423 | $0.042 | $0.084 |
| Mistral: Ministral 3 3B 2512 | $0.1 | $0.1 |
| OpenAI: gpt-oss-120b | $0.037 | $0.17 |
| OpenAI: gpt-oss-20b | $0.018 | $0.09 |
| Qwen: Qwen3 30B A3B Instruct 2507 | $0.048 | $0.193 |
Other IBM models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Granite 4.0 Micro | $0.017 | $0.112 |