Llama 3.3 Nemotron Super 49B API pricing on Runcrate
Every Runcrate Llama 3.3 Nemotron Super 49B rate we track, compared against 1 provider serving the same model.
Runcrate Llama 3.3 Nemotron Super 49B rates
Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.
| Mode | Input | Output |
|---|---|---|
| Standard | $0.400/1M | $0.400/1M |
About Llama 3.3 Nemotron Super 49B
- Creator
- NVIDIA
- Context
- 131K
- Modalities
- text
About Runcrate
Runcrate provides a cloud-based platform for inference and compute services, focused on scaling AI workflows.
Frequently Asked Questions
How much does Llama 3.3 Nemotron Super 49B cost on Runcrate?
Runcrate serves Llama 3.3 Nemotron Super 49B from $0.400 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.
Is Runcrate the cheapest way to run Llama 3.3 Nemotron Super 49B?
Of the 1 providers we track serving Llama 3.3 Nemotron Super 49B, Runcrate currently has the lowest input price. Rates change often, and throughput and latency differ between providers — check the comparison above before committing.
Does Runcrate offer batch or cached pricing for Llama 3.3 Nemotron Super 49B?
We track 1 Llama 3.3 Nemotron Super 49B offering from Runcrate. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.