Nemotron API pricing
Every Nemotron model we track, with the cheapest price per 1M tokens across 11 providers. Collected daily.
Nemotron API prices by model
Standard (real-time) rates per 1M tokens, newest first. “Cheapest” is the lowest input price across every provider we track; batch and cached-input rates are on each model page.
| Model | Cheapest input | Output | Cheapest on |
|---|---|---|---|
| Nemotron 3.5 Lightning | $0.047 | $0.134 | OpenRouter |
| Llama 3.1 Nemotron 70B131K context | $0.880 | $0.880 | Together AI |
| Nemotron 3 Nano 30B262K context | $0.050 | $0.200 | Crusoe |
| Nemotron 3 Ultra 550B A55B | $0.500 | $2.20 | OpenRouter |
| Nemotron 3 Super 120B262K context | $0.080 | $0.450 | OpenRouter |
| Nemotron 3 Nano Omni 30B A3B Reasoning | $0.200 | $0.800 | Runcrate |
| Nemotron Nano 12B v2 VL131K context | $0.200 | $0.600 | Amazon AWS |
| Nemotron Nano 9B v2131K context | $0.060 | $0.230 | Amazon AWS |
| NVIDIA Nemotron Nano 9B V2 | $0.060 | $0.250 | Together AI |
Nemotron price history
Average input price per 1M tokens across the providers serving each model, last 90 days.
Frequently Asked Questions
What is the cheapest Nemotron API?
Of the 9 Nemotron models we track, Nemotron 3.5 Lightning has the lowest input price right now: $0.047 per 1M input tokens on OpenRouter. Prices are collected daily from 11 providers — see the table above for every model.
Where is Nemotron 3.5 Lightning cheapest?
Nemotron 3.5 Lightning is served by 4 providers we track. The lowest standard input price is $0.047 per 1M tokens on OpenRouter. Throughput, latency and context limits differ between hosts, so check the model page before committing.