Crusoe
Energy-first AI cloud powered by renewable and low-carbon energy
Last reviewed Mar 14, 2026
Crusoe Cloud offers energy-first, high-performance GPU compute (H100, H200, A100, L40S, B200, GB200, and AMD MI300X/MI355X) alongside managed inference and fine-tuning services, backed by vertically integrated AI data centers.
Available GPUs
Hourly on-demand pricing. Click column headers to sort.
Prices last updated: August 12, 2026
Crusoe pricing by GPU
Configurations, price rank, and alternatives for one GPU at a time.
LLM API Pricing
Pay-per-token pricing. Prices shown per 1M tokens.
Prices last updated: August 10, 2026
| Model | Input/1M | Output/1M | |||
|---|---|---|---|---|---|
| $0.050 | $0.200 | ||||
| $0.050 | $0.200 | ||||
| $0.140 | $0.280 | ||||
| $0.140 | $0.400 | ||||
| $0.220 | $0.800 | ||||
| $0.250 | $0.750 | ||||
| $0.300 | $1.83 | ||||
| $0.500 | $1.50 | ||||
| $0.700 | $3.50 | ||||
| $1.00 | $3.20 | ||||
Crusoe pricing by model
Input, output, and batch rates, plus alternatives, for one model at a time.
Pros & Cons
Advantages
- Great GPU availability for on-demand use
- Published hourly on-demand rates for most GPU types
- Renewable and low-carbon energy sourcing with vertical integration from power to cloud
- 99.5% uptime commitment with 24/7 enterprise support
- Managed inference and serverless fine-tuning available alongside raw GPU instances
- Fast instance deployment
Limitations
- Contact sales required for newest GPU models (GB200, B200, MI355X)
- Spot capacity is quoted by sales rather than published as a self-service rate
- Data center footprint concentrated in the United States with limited European capacity
Key Features
Energy-First Infrastructure
Vertically integrated power sourcing across wind, solar, hydropower, geothermal, natural gas, and carbon capture
Fast Spin Up
Machines are deployed quickly when they are launched
Persistent Storage
Persistent and shared disks, object storage, and a container registry for AI datasets
SXM Support
Provides instances with NVIDIA SXM GPU interconnects
SLA Guarantee
99.5% uptime commitment backed by 24/7 enterprise-grade support
Managed Inference
LLM inference with up to 9.9x faster time-to-first-token and 5x higher throughput versus vLLM
Serverless Fine-Tuning
LoRA-based supervised fine-tuning of open models through Crusoe Intelligence Foundry, with no cluster provisioning
Cluster Orchestration
Managed Kubernetes, Managed Slurm, and fault-tolerant AutoClusters for multi-node training
Command Center
Unified operations platform for high-performance AI workloads
Security Certifications
ISO 27001 and ISO 42001 certified for information security and responsible AI governance
Compute Services
GPU Instances
High-performance GPU instances powered by NVIDIA GPUs
CPU Instances
General-purpose and storage-optimized CPU instances for data processing
Managed Inference
Fully managed LLM inference spanning serverless, self-serve dedicated, and tailored deployments
- Open models including DeepSeek V4, GLM 5.2, Kimi K2.6, Llama 3.3, Gemma 4, Qwen3, GPT-OSS, and Nemotron 3
- Serverless Inference billed per million input, output, and cached tokens
- Self-Serve Deployments give dedicated H100 or H200 endpoints without sales engagement
- Tailored Deployments are benchmarked and optimized with Crusoe's team
- Provisioned throughput with AI Model Units (AMUs)
- Up to 9.9x faster time-to-first-token and 5x higher throughput versus vLLM
Serverless Fine-Tuning
LoRA-based supervised fine-tuning of open models in Crusoe Intelligence Foundry
- No cluster provisioning required
- Token-based pricing tiered by base model parameter count
- One-click deployment of the tuned model to an inference endpoint
- Full portability of the resulting weights
Managed Kubernetes and Slurm
Managed Kubernetes (CMK) and Managed Slurm clusters, plus fault-tolerant AutoClusters, for AI workloads
- Simplifies deployment and scaling
- Supports both GPU and CPU resources
- Billed per cluster hour on top of instance costs
- Container registry integration
Reserved Clusters
Large-scale reserved GPU clusters for enterprise and research workloads
- Custom node configurations
- Flexible interconnect options
- Long-term reservation for consistent availability
Crusoe Edge Zones
High-performance AI infrastructure powered by Crusoe Spark modular AI data centers
- Scalable edge computing capabilities
- Modular AI data center deployment
- Reduced latency for edge AI workloads
Pricing Options
| Option | Details |
|---|---|
| On-Demand Instances | Pay by the hour for maximum agility and unthrottled compute |
| Spot Instances | Use spare capacity at significant discounts compared to on-demand pricing |
| Reserved Instances | Lock in guaranteed resources at lowest rates with custom commitments |
| Pay-as-you-go Inference | Flexible pricing for LLM inference based on token usage |
| Provisioned Throughput | Guaranteed inference capacity with AI Model Units, priced lower for longer commitments |
| Dedicated Inference Deployments | Hourly per-GPU pricing for self-serve dedicated endpoints, with monthly and volume rates via sales |
| Serverless Fine-Tuning | Per-token pricing tiered by the parameter count of the base model being tuned |
Availability & Support
Regions
Data centers sited near low-cost energy in the United States, including a large-scale campus in Abilene, Texas, plus European capacity in Iceland and Norway
Support
Documentation, learning hub, and 24/7 enterprise-grade support with a 99.5% uptime commitment; SOC 2, ISO 27001, and ISO 42001 certifications
Getting Started
- 1
Create an account
Sign up for Crusoe Cloud services through their website
- 2
Select GPU instance
Choose from available GPU types including L40S, A100, H100, H200, B200, GB200, and AMD options
- 3
Configure storage
Set up persistent storage options for your workloads
- 4
Launch instance
Deploy via the console, or provision programmatically with the Crusoe CLI or Terraform
Compare Providers
Find the best prices for the same GPUs and models from other providers