Skip to main content
Groq logo

Llama 3.3 70B API pricing on Groq

Every Groq Llama 3.3 70B rate we track, compared against 8 providers serving the same model.

Input from /1M
$0.590
Output /1M
$0.790
Price rank
#5 of 8
Last updated
August 3, 2026

The cheapest Llama 3.3 70B input price we currently track is $0.100/1M on Deep Infra.

Groq Llama 3.3 70B rates

Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.

Groq Llama 3.3 70B pricing by mode
ModeInputOutput
Standard$0.590/1M$0.790/1M

About Llama 3.3 70B

Creator
Meta
Context
128K
Modalities
text
Knowledge cutoff
Mar 2024
Tool calling
Yes
Open weights
Yes
Full Llama 3.3 70B details and every provider →

Llama 3.3 70B on other providers

About Groq

Groq provides ultra-fast LLM inference powered by their custom LPU (Language Processing Unit) hardware, offering the fastest token generation speeds in the industry.

Frequently Asked Questions

How much does Llama 3.3 70B cost on Groq?

Groq serves Llama 3.3 70B from $0.590 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.

Is Groq the cheapest way to run Llama 3.3 70B?

Groq ranks #5 of 8 providers we track serving Llama 3.3 70B. The cheapest input price right now is on Deep Infra. Price is only one factor — throughput, latency, and context limits differ between providers.

Does Groq offer batch or cached pricing for Llama 3.3 70B?

We track 1 Llama 3.3 70B offering from Groq. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.