Google · live price
Gemma 4 26B A4B API pricing
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
- Input
- $0.09
- per 1M tokens
- Output
- $0.3
- per 1M tokens
- Cached input
- $0.05
- per 1M tokens
- Context
- 262K
- 236K max output
List price from OpenRouter, which passes Google's price through without markup. Refreshed hourly. Released April 3, 2026. Model id google/gemma-4-26b-a4b-it.
What Gemma 4 26B A4B costs in practice
1,000 chat messages
$0.24
1,000 input + 500 output tokens each
1,000 RAG questions
$0.87
8,000 input + 500 output tokens each
100 coding-agent steps
$0.33
30,000 input + 2,000 output tokens each, no caching
Gemma 4 26B A4B price history
Gemma 4 26B A4B price by host
- Darkbloom$0.042 in · $0.22 out · 131K ctx · 100.0% uptime 24h
- DekaLLM$0.06 in · $0.33 out · 262K ctx · 99.7% uptime 24h
- NextBit$0.09 in · $0.3 out · 262K ctx · 99.9% uptime 24h
- CoreWeave$0.1 in · $0.3 out · 262K ctx · 100.0% uptime 24h
- Cloudflare$0.1 in · $0.3 out · 256K ctx · 99.7% uptime 24h
- Makora$0.08 in · $0.32 out · 256K ctx · 98.7% uptime 24h
- DeepInfra$0.07 in · $0.34 out · 262K ctx · 99.3% uptime 24h
- Venice$0.13 in · $0.4 out · 256K ctx · 99.5% uptime 24h
- Parasail$0.13 in · $0.4 out · 262K ctx · 99.7% uptime 24h
- Novita$0.13 in · $0.4 out · 262K ctx · 99.6% uptime 24h
- SiliconFlow$0.14 in · $0.4 out · 262K ctx · 94.9% uptime 24h
- Io Net$0.15 in · $0.5 out · 262K ctx · 97.7% uptime 24h
- Google$0.15 in · $0.6 out · 262K ctx · 96.4% uptime 24h
Where to buy Google credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Gemma 4 26B A4B cost calculator
Cheaper alternatives to Gemma 4 26B A4B
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Qwen: Qwen3.7 Flash | $0.03 | $0.13 |
| DeepSeek: DeepSeek V4 Flash 0423 | $0.042 | $0.084 |
| Qwen: Qwen3.5-9B | $0.1 | $0.15 |
| Qwen: Qwen3.5-Flash | $0.065 | $0.26 |
| Mistral: Ministral 3 3B 2512 | $0.1 | $0.1 |
| OpenAI: gpt-oss-safeguard-20b | $0.075 | $0.3 |
Other Google models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Gemma 4 31B | $0.09 | $0.34 |
| Gemini 3.1 Flash Lite Preview | $0.25 | $1.50 |
| Gemini 3.1 Flash Lite | $0.25 | $1.50 |
| Nano Banana 2 (Gemini 3.1 Flash Image Preview) | $0.5 | $3 |
| Gemini 3.1 Pro Preview Custom Tools | $2 | $12 |
| Gemini 3.1 Pro Preview | $2 | $12 |