Skip to main content
Heabsy logo

Llama 3.1 8B API pricing on Heabsy

Every Heabsy Llama 3.1 8B rate we track, compared against 6 providers serving the same model.

Input from /1M
$0.080
Output /1M
$0.120
Price rank
#3 of 6
Last updated
September 30, 2026

The cheapest Llama 3.1 8B input price we currently track is $0.020/1M on Novita AI.

Heabsy Llama 3.1 8B rates

Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.

Heabsy Llama 3.1 8B pricing by mode
ModeInputOutput
Standard$0.080/1M$0.120/1M

About Llama 3.1 8B

Creator
Meta
Context
128K
Modalities
text
Knowledge cutoff
Dec 2023
Tool calling
Yes
Open weights
Yes
Full Llama 3.1 8B details and every provider →

Llama 3.1 8B on other providers

About Heabsy

Heabsy is an OpenAI- and Anthropic-compatible inference API for open-weight models, operated by FEYA, s.r.o. in Bratislava, Slovakia, an EU company with no US parent. Models in the EEA tier, including Qwen3.8 27B, run on GPUs that Heabsy owns and operates inside the European Economic Area, with zero data retention: prompts and completions are processed in memory and never written to logs. A routed tier adds more than 25 open models (DeepSeek, Kimi, GLM, gpt-oss, Llama, Gemma, Qwen and others) bought from global providers under a European contract; every model in the catalog is labelled with where it is processed. One API key, one EU invoice, per-token pricing with no minimum spend or subscription.

Frequently Asked Questions

How much does Llama 3.1 8B cost on Heabsy?

Heabsy serves Llama 3.1 8B from $0.080 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.

Is Heabsy the cheapest way to run Llama 3.1 8B?

Heabsy ranks #3 of 6 providers we track serving Llama 3.1 8B. The cheapest input price right now is on Novita AI. Price is only one factor — throughput, latency, and context limits differ between providers.

Does Heabsy offer batch or cached pricing for Llama 3.1 8B?

We track 1 Llama 3.1 8B offering from Heabsy. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.