Skip to main content
Lyceum logo

Llama 3.1 Nemotron Ultra 253B API pricing on Lyceum

Every Lyceum Llama 3.1 Nemotron Ultra 253B rate we track, compared against 1 provider serving the same model.

Input from /1M
$0.600
Output /1M
$1.80
Price rank
#1 of 1
Last updated
August 13, 2026

Lyceum Llama 3.1 Nemotron Ultra 253B rates

Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.

Lyceum Llama 3.1 Nemotron Ultra 253B pricing by mode
ModeInputOutput
Standard$0.600/1M$1.80/1M

About Llama 3.1 Nemotron Ultra 253B

Creator
NVIDIA
Context
131K
Modalities
text
Full Llama 3.1 Nemotron Ultra 253B details and every provider →

About Lyceum

Lyceum offers serverless inference, dedicated endpoints, GPU VMs, and large-scale clusters for AI workloads. The platform is designed to facilitate the deployment and management of AI solutions without the complexity of traditional infrastructure.

Frequently Asked Questions

How much does Llama 3.1 Nemotron Ultra 253B cost on Lyceum?

Lyceum serves Llama 3.1 Nemotron Ultra 253B from $0.600 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.

Is Lyceum the cheapest way to run Llama 3.1 Nemotron Ultra 253B?

Of the 1 providers we track serving Llama 3.1 Nemotron Ultra 253B, Lyceum currently has the lowest input price. Rates change often, and throughput and latency differ between providers — check the comparison above before committing.

Does Lyceum offer batch or cached pricing for Llama 3.1 Nemotron Ultra 253B?

We track 1 Llama 3.1 Nemotron Ultra 253B offering from Lyceum. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.