NVIDIA · live price
Nemotron 3 Super API pricing
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
- Input
- $0.08
- per 1M tokens
- Output
- $0.45
- per 1M tokens
- Cached input
- –
- per 1M tokens
- Context
- 262K
- 236K max output
List price from OpenRouter, which passes NVIDIA's price through without markup. Refreshed hourly. Released March 11, 2026. Model id nvidia/nemotron-3-super-120b-a12b.
What Nemotron 3 Super costs in practice
1,000 chat messages
$0.31
1,000 input + 500 output tokens each
1,000 RAG questions
$0.87
8,000 input + 500 output tokens each
100 coding-agent steps
$0.33
30,000 input + 2,000 output tokens each, no caching
Nemotron 3 Super price history
Nemotron 3 Super price by host
- DeepInfra$0.085 in · $0.4 out · 262K ctx · 100.0% uptime 24h
- DekaLLM$0.08 in · $0.45 out · 262K ctx · 99.1% uptime 24h
Where to buy NVIDIA credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Nemotron 3 Super cost calculator
Cheaper alternatives to Nemotron 3 Super
| Model | Input / 1M | Output / 1M |
|---|---|---|
| DeepSeek: DeepSeek V4.1 Flash | $0.03 | $0.5 |
| Qwen: Qwen3.7 Flash | $0.03 | $0.13 |
| DeepSeek: DeepSeek V4 Flash 0423 | $0.042 | $0.084 |
| Google: Gemma 4 26B A4B | $0.09 | $0.3 |
| Google: Gemma 4 31B | $0.09 | $0.34 |
| Qwen: Qwen3.5-9B | $0.1 | $0.15 |
Other NVIDIA models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Nemotron 3 Ultra | $0.6 | $2.40 |
| Nemotron 3.5 Content Safety | $0.2 | $0.2 |
| Nemotron 3 Nano 30B A3B | $0.05 | $0.2 |
| Nemotron 3.5 Lightning | $0.059 | $0.17 |