Llama 3.3 70B API pricing on Groq
Every Groq Llama 3.3 70B rate we track, compared against 8 providers serving the same model.
The cheapest Llama 3.3 70B input price we currently track is $0.100/1M on Deep Infra.
Groq Llama 3.3 70B rates
Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.
| Mode | Input | Output |
|---|---|---|
| Standard | $0.590/1M | $0.790/1M |
About Llama 3.3 70B
- Creator
- Meta
- Context
- 128K
- Modalities
- text
- Knowledge cutoff
- Mar 2024
- Tool calling
- Yes
- Open weights
- Yes
Llama 3.3 70B on other providers
About Groq
Groq provides ultra-fast LLM inference powered by their custom LPU (Language Processing Unit) hardware, offering the fastest token generation speeds in the industry.
Frequently Asked Questions
How much does Llama 3.3 70B cost on Groq?
Groq serves Llama 3.3 70B from $0.590 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.
Is Groq the cheapest way to run Llama 3.3 70B?
Groq ranks #5 of 8 providers we track serving Llama 3.3 70B. The cheapest input price right now is on Deep Infra. Price is only one factor — throughput, latency, and context limits differ between providers.
Does Groq offer batch or cached pricing for Llama 3.3 70B?
We track 1 Llama 3.3 70B offering from Groq. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.