Loading Comparison
Fetching pricing data and provider information...
Loading Comparison
Fetching pricing data and provider information...
Compare GPU and LLM inference API pricing between Lyceum and Wafer. Find the best rates for AI training, inference, and ML workloads.
Provider 1
Provider 2
| GPU Model ↑ | Lyceum Price | Wafer Price | Price Diff ↕ | Sources |
|---|---|---|---|---|
A100 SXM 80GB VRAM • Lyceum | Not Available | — | ||
A100 SXM 80GB VRAM • | ||||
B200 180GB VRAM • Lyceum | Not Available | — | ||
B200 180GB VRAM • | ||||
H100 SXM 80GB VRAM • Lyceum | Not Available | — | ||
H100 SXM 80GB VRAM • | ||||
H200 141GB VRAM • Lyceum | Not Available | — | ||
H200 141GB VRAM • | ||||
HGX B300 288GB VRAM • Lyceum | Not Available | — | ||
HGX B300 288GB VRAM • | ||||
L40S 48GB VRAM • Lyceum | Not Available | — | ||
L40S 48GB VRAM • | ||||
A100 SXM 80GB VRAM • Lyceum | Not Available | — | ||
A100 SXM 80GB VRAM • | ||||
B200 180GB VRAM • Lyceum | Not Available | — | ||
B200 180GB VRAM • | ||||
H100 SXM 80GB VRAM • Lyceum | Not Available | — | ||
H100 SXM 80GB VRAM • | ||||
H200 141GB VRAM • Lyceum | Not Available | — | ||
H200 141GB VRAM • | ||||
HGX B300 288GB VRAM • Lyceum | Not Available | — | ||
HGX B300 288GB VRAM • | ||||
L40S 48GB VRAM • Lyceum | Not Available | — | ||
L40S 48GB VRAM • | ||||
| Model ↑ | Lyceum | Wafer | Input Diff ↕ |
|---|---|---|---|
DeepSeek | $0.300 in $0.450 out | Not available | — |
DeepSeek | $0.150 in $0.300 out | Not available | — |
DeepSeek | $1.75 in $3.50 out | Not available | — |
Google | $0.100 in $0.300 out | Not available | — |
Zhipu | $1.00 in $3.20 out | Not available | — |
Zhipu | $1.40 in $4.40 out | $1.00 in $3.20 out | $0.400 |
Zhipu | $1.50 in $4.50 out | $1.20 in $4.10 out | $0.300 |
OpenAI | $0.150 in $0.600 out | Not available | — |
Moonshot | $1.00 in $4.00 out | Not available | — |
Moonshot | $0.500 in $2.50 out | Not available | — |
Moonshot | $3.00 in $15.00 out | Not available | — |
NVIDIA | $0.600 in $1.80 out | Not available | — |
MiniMax | $0.400 in $2.00 out | Not available | — |
NVIDIA | $0.060 in $0.240 out | Not available | — |
NVIDIA | $0.060 in $0.240 out | Not available | — |
Explore how these providers compare to other popular GPU cloud services
Compare Lyceum with another leading provider
Compare Lyceum with another leading provider
Compare Lyceum with another leading provider
Compare Lyceum with another leading provider
Compare Lyceum with another leading provider
Compare Lyceum with another leading provider
API access to open-source models with pay-per-token pricing.
Reserved GPU capacity for production models ensuring low latency and high availability.
Run training jobs on GPUs without the need for infrastructure management.
Full root access to customizable GPU instances ready in seconds.
Support for configurations from 8 to 8,000 GPUs with InfiniBand connectivity.
Pay-as-you-go API access to hosted open-source models including GLM, Kimi, Qwen, and DeepSeek with no infrastructure management
Custom-tuned inference deployments with performance guarantees, provisioned in under 24 hours
Agents profile inference bottlenecks and tune across serving engines (vLLM, SGLang, TensorRT-LLM), custom kernels (CUDA, HIP, Triton, NKI), quantization (FP8/FP4), and decode strategies
Workloads run on NVIDIA B200/B300, AMD MI350X/MI355X, and AWS Trainium depending on the model and traffic shape
OpenAI-compatible endpoint at pass.wafer.ai/v1 and Anthropic-compatible endpoint at pass.wafer.ai/v1/messages, both using Bearer token authentication
Cached input tokens are billed at reduced rates on supported models
Prepaid credits with separate input and output token rates per model and no subscription
Reduced rates for cached input tokens on supported models
Custom pricing for dedicated deployments with tuned performance targets, arranged with the sales team
Create an account on the Lyceum platform.
Select between inference, training, or virtual machine options.
Use provided API or CLI instructions to submit your workload.
Utilize dashboard features to track resource usage and performance.
Contact the sales or support team for assistance.
Sign up at app.wafer.ai and load credits for pay-as-you-go usage
Create a key in the console and pass it as a Bearer token
Call the OpenAI-compatible endpoint at pass.wafer.ai/v1 with a model from the serverless catalog
GPUs hosted in European data centres
Contact support via sales inquiry or demo booking.
Documentation at docs.wafer.ai, email support (hi@wafer.ai), and scheduled onboarding calls for enterprise