Skip to main content
Hugging Face logo

L4 pricing on Hugging Face

Every Hugging Face L4 offering we track, normalized to a per-GPU hourly rate and ranked against 22 providers with current L4 pricing.

From / GPU-hr
$0.700
Price rank
#12 of 22
Configurations
4
Last updated
September 28, 2026

The cheapest L4 we currently track is $0.321/hr on Vast.ai.

Hugging Face L4 rates

Hourly prices per GPU. Multi-GPU rows show the per-GPU rate, not the instance total.

Hugging Face L4 pricing by configuration
Price / GPU-hrTypeGPUs
$0.700On-demand1
$0.800On-demand1
$0.950On-demand4
$0.950On-demand4

L4 specifications

VRAM
24GB
Architecture
Ada Lovelace
CUDA Cores
7,424
Tensor Cores
232
Memory Bandwidth
300 GB/s
TDP
72W
Full L4 specs and every provider →

L4 on other providers

About Hugging Face

Hugging Face runs the Hub for open models and datasets. Inference Endpoints deploys a model on dedicated GPUs on AWS or Google Cloud, billed per hour of instance time, and Inference Providers routes API requests to partner inference providers at those providers' own rates.

Frequently Asked Questions

How much does the L4 cost on Hugging Face?

Hugging Face lists the L4 from $0.700 per GPU-hour across 4 tracked configurations. Prices are collected daily — see the table above for the current rates.

Is Hugging Face the cheapest place to rent the L4?

Hugging Face ranks #12 of 22 providers we track with current L4 pricing. The cheapest right now is Vast.ai. Price is only one factor — region availability, network, and storage costs differ between providers.

Which L4 configurations does Hugging Face offer?

We currently track 4 L4 offerings from Hugging Face, covering the GPU counts, pricing types, and regions listed in the table above. Every price is normalized to a per-GPU hourly rate so it can be compared across providers.