NVIDIA
Nemotron 3 Nano Omni 30B A3B Reasoning
Nemotron 3 Nano Omni 30B A3B Reasoning by NVIDIA — compare inference API pricing across providers.
Input from
$0.200 / 1M tokens
across 1 provider
API Pricing
| Provider | Input / 1M | Output / 1M |
|---|---|---|
| $0.200 | $0.800 |
Prices updated daily. Last check: Jul 16, 2026
Performance & Benchmarks
Source: Artificial Analysis →Intelligence
14.9 / 100
Output Speed
318 t/s
Latency (TTFT)
543ms
Reasoning & Knowledge
- GPQA Diamond46.9%
- Humanity's Last Exam5.3%
Coding
- SciCode27.8%
Agentic & Tool Use
- Terminal-Bench Hard8.3%
- τ²-bench45.3%
Instruction & Long Context
- IFBench63.2%
- Long-Context Reasoning35.7%
Benchmarks measured Jul 2026. Scores are independent evaluations, not vendor-reported.
Model Details
General
- Creator
- NVIDIA
- Modalities
- Text
Capabilities
- Tool Calling
- No
- Open Source
- No