Llama 3.1 Nemotron Ultra 253B API pricing on Lyceum
Every Lyceum Llama 3.1 Nemotron Ultra 253B rate we track, compared against 1 provider serving the same model.
Lyceum Llama 3.1 Nemotron Ultra 253B rates
Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.
| Mode | Input | Output |
|---|---|---|
| Standard | $0.600/1M | $1.80/1M |
About Llama 3.1 Nemotron Ultra 253B
- Creator
- NVIDIA
- Context
- 131K
- Modalities
- text
About Lyceum
Lyceum offers serverless inference, dedicated endpoints, GPU VMs, and large-scale clusters for AI workloads. The platform is designed to facilitate the deployment and management of AI solutions without the complexity of traditional infrastructure.
Frequently Asked Questions
How much does Llama 3.1 Nemotron Ultra 253B cost on Lyceum?
Lyceum serves Llama 3.1 Nemotron Ultra 253B from $0.600 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.
Is Lyceum the cheapest way to run Llama 3.1 Nemotron Ultra 253B?
Of the 1 providers we track serving Llama 3.1 Nemotron Ultra 253B, Lyceum currently has the lowest input price. Rates change often, and throughput and latency differ between providers — check the comparison above before committing.
Does Lyceum offer batch or cached pricing for Llama 3.1 Nemotron Ultra 253B?
We track 1 Llama 3.1 Nemotron Ultra 253B offering from Lyceum. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.