Skip to main content
NVIDIA

Nemotron 3 Nano Omni 30B A3B Reasoning

Nemotron 3 Nano Omni 30B A3B Reasoning by NVIDIA — compare inference API pricing across providers.

Input from
$0.200 / 1M tokens
across 1 provider

API Pricing

ProviderInput / 1MOutput / 1M
$0.200$0.800

Prices updated daily. Last check: Jul 16, 2026

Performance & Benchmarks

Source: Artificial Analysis →
Intelligence
14.9 / 100
Output Speed
318 t/s
Latency (TTFT)
543ms

Reasoning & Knowledge

  • GPQA Diamond46.9%
  • Humanity's Last Exam5.3%

Coding

  • SciCode27.8%

Agentic & Tool Use

  • Terminal-Bench Hard8.3%
  • τ²-bench45.3%

Instruction & Long Context

  • IFBench63.2%
  • Long-Context Reasoning35.7%

Benchmarks measured Jul 2026. Scores are independent evaluations, not vendor-reported.

Model Details

General

Creator
NVIDIA
Modalities
Text

Capabilities

Tool Calling
No
Open Source
No