Energy-first AI cloud powered by renewable and low-carbon energy
Last reviewed Mar 14, 2026
Crusoe Cloud offers energy-first, high-performance GPU compute (H100, H200, A100, L40S, B200, GB200, and AMD MI300X/MI355X) alongside managed inference and fine-tuning services, backed by vertically integrated AI data centers.
Hourly on-demand pricing. Click column headers to sort.
Prices last updated: September 17, 2026
Configurations, price rank, and alternatives for one GPU at a time.
Pay-per-token pricing. Prices shown per 1M tokens.
Prices last updated: September 19, 2026
| Model | Input/1M | Output/1M | |||
|---|---|---|---|---|---|
| $0.050 | $0.200 | ||||
| $0.050 | $0.200 | ||||
| $0.050 | $0.200 | ||||
| $0.140 | $0.280 | ||||
| $0.140 | $0.400 | ||||
| $0.150 | $0.500 | ||||
| $0.300 | $1.83 | ||||
| $0.300 | $2.40 | ||||
| $0.700 | $3.50 | ||||
| $1.00 | $3.20 | ||||
Input, output, and batch rates, plus alternatives, for one model at a time.
Vertically integrated power sourcing across wind, solar, hydropower, geothermal, natural gas, and carbon capture
Machines are deployed quickly when they are launched
Persistent and shared disks, object storage, and a container registry for AI datasets
Provides instances with NVIDIA SXM GPU interconnects
99.5% uptime commitment backed by 24/7 enterprise-grade support
LLM inference with up to 9.9x faster time-to-first-token and 5x higher throughput versus vLLM
LoRA-based supervised fine-tuning of open models through Crusoe Intelligence Foundry, with no cluster provisioning
Managed Kubernetes, Managed Slurm, and fault-tolerant AutoClusters for multi-node training
Unified operations platform for high-performance AI workloads
ISO 27001 and ISO 42001 certified for information security and responsible AI governance
High-performance GPU instances powered by NVIDIA GPUs
General-purpose and storage-optimized CPU instances for data processing
Fully managed LLM inference spanning serverless, self-serve dedicated, and tailored deployments
LoRA-based supervised fine-tuning of open models in Crusoe Intelligence Foundry
Managed Kubernetes (CMK) and Managed Slurm clusters, plus fault-tolerant AutoClusters, for AI workloads
Large-scale reserved GPU clusters for enterprise and research workloads
High-performance AI infrastructure powered by Crusoe Spark modular AI data centers
| Option | Details |
|---|---|
| On-Demand Instances | Pay by the hour for maximum agility and unthrottled compute |
| Spot Instances | Use spare capacity at significant discounts compared to on-demand pricing |
| Reserved Instances | Lock in guaranteed resources at lowest rates with custom commitments |
| Pay-as-you-go Inference | Flexible pricing for LLM inference based on token usage |
| Provisioned Throughput | Guaranteed inference capacity with AI Model Units, priced lower for longer commitments |
| Dedicated Inference Deployments | Hourly per-GPU pricing for self-serve dedicated endpoints, with monthly and volume rates via sales |
| Serverless Fine-Tuning | Per-token pricing tiered by the parameter count of the base model being tuned |
Data centers sited near low-cost energy in the United States, including a large-scale campus in Abilene, Texas, plus European capacity in Iceland and Norway
Documentation, learning hub, and 24/7 enterprise-grade support with a 99.5% uptime commitment; SOC 2, ISO 27001, and ISO 42001 certifications
Sign up for Crusoe Cloud services through their website
Choose from available GPU types including L40S, A100, H100, H200, B200, GB200, and AMD options
Set up persistent storage options for your workloads
Deploy via the console, or provision programmatically with the Crusoe CLI or Terraform
Crusoe offers various GPU types including A100 PCIE, A100 SXM, H100 SXM, H200, L40S, MI300X. Check the pricing table above for current availability and pricing.
Create an account, Select GPU instance, Configure storage, Launch instance
Crusoe's main advantages include: Great GPU availability for on-demand use, Published hourly on-demand rates for most GPU types, Renewable and low-carbon energy sourcing with vertical integration from power to cloud, 99.5% uptime commitment with 24/7 enterprise support, Managed inference and serverless fine-tuning available alongside raw GPU instances, Fast instance deployment.
Crusoe's main limitations include: Contact sales required for newest GPU models (GB200, B200, MI355X), Spot capacity is quoted by sales rather than published as a self-service rate, Data center footprint concentrated in the United States with limited European capacity.
Find the best prices for the same GPUs and models from other providers