Skip to main content
NVIDIA

Nemotron API pricing

Every Nemotron model we track, with the cheapest price per 1M tokens across 11 providers. Collected daily.

Cheapest input /1M
$0.047
Nemotron 3.5 Lightning
Models
9
Providers
11
Last updated
October 9, 2026

Nemotron API prices by model

Standard (real-time) rates per 1M tokens, newest first. “Cheapest” is the lowest input price across every provider we track; batch and cached-input rates are on each model page.

Nemotron API pricing by model
ModelCheapest inputOutputCheapest on
Nemotron 3.5 Lightning$0.047$0.134OpenRouter
Llama 3.1 Nemotron 70B131K context$0.880$0.880Together AI
Nemotron 3 Nano 30B262K context$0.050$0.200Crusoe
Nemotron 3 Ultra 550B A55B$0.500$2.20OpenRouter
Nemotron 3 Super 120B262K context$0.080$0.450OpenRouter
Nemotron 3 Nano Omni 30B A3B Reasoning$0.200$0.800Runcrate
Nemotron Nano 12B v2 VL131K context$0.200$0.600Amazon AWS
Nemotron Nano 9B v2131K context$0.060$0.230Amazon AWS
NVIDIA Nemotron Nano 9B V2$0.060$0.250Together AI

Nemotron price history

Average input price per 1M tokens across the providers serving each model, last 90 days.

Frequently Asked Questions

What is the cheapest Nemotron API?

Of the 9 Nemotron models we track, Nemotron 3.5 Lightning has the lowest input price right now: $0.047 per 1M input tokens on OpenRouter. Prices are collected daily from 11 providers — see the table above for every model.

Where is Nemotron 3.5 Lightning cheapest?

Nemotron 3.5 Lightning is served by 4 providers we track. The lowest standard input price is $0.047 per 1M tokens on OpenRouter. Throughput, latency and context limits differ between hosts, so check the model page before committing.

Other model families