NVIDIA · live price
Nemotron 3 Nano 30B A3B API pricing
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
- Input
- $0.05
- per 1M tokens
- Output
- $0.2
- per 1M tokens
- Cached input
- $0.03
- per 1M tokens
- Context
- 262K
- 236K max output
List price from OpenRouter, which passes NVIDIA's price through without markup. Refreshed hourly. Released December 14, 2025. Model id nvidia/nemotron-3-nano-30b-a3b.
What Nemotron 3 Nano 30B A3B costs in practice
1,000 chat messages
$0.15
1,000 input + 500 output tokens each
1,000 RAG questions
$0.50
8,000 input + 500 output tokens each
100 coding-agent steps
$0.19
30,000 input + 2,000 output tokens each, no caching
Nemotron 3 Nano 30B A3B price history
Nemotron 3 Nano 30B A3B price by host
- Crusoe$0.05 in · $0.2 out · 262K ctx · 100.0% uptime 24h
- Novita$0.05 in · $0.2 out · 262K ctx · 99.6% uptime 24h
- DeepInfra$0.05 in · $0.2 out · 262K ctx · 99.9% uptime 24h
- Nebius$0.06 in · $0.24 out · 262K ctx · 94.4% uptime 24h
Where to buy NVIDIA credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Nemotron 3 Nano 30B A3B cost calculator
Cheaper alternatives to Nemotron 3 Nano 30B A3B
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Qwen: Qwen3.7 Flash | $0.03 | $0.13 |
| DeepSeek: DeepSeek V4 Flash 0423 | $0.042 | $0.084 |
| OpenAI: gpt-oss-120b | $0.037 | $0.17 |
| OpenAI: gpt-oss-20b | $0.018 | $0.09 |
| Qwen: Qwen3 30B A3B Instruct 2507 | $0.048 | $0.193 |
| Google: Gemma 3 4B | $0.05 | $0.1 |
Other NVIDIA models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Nemotron 3 Super | $0.08 | $0.45 |
| Nemotron 3 Ultra | $0.6 | $2.40 |
| Nemotron 3.5 Content Safety | $0.2 | $0.2 |
| Nemotron 3.5 Lightning | $0.059 | $0.17 |