Loading Comparison
Fetching pricing data and provider information...
Compare GPU and LLM inference API pricing between Sesterce and Together AI. Find the best rates for AI training, inference, and ML workloads.
Provider 1
Provider 2
Average Price Difference: $1.49/hour between comparable GPUs
| GPU Model ↑ | Sesterce Price | Together AI Price | Price Diff ↕ | Sources |
|---|---|---|---|---|
A10 24GB VRAM • Sesterce | Not Available | — | ||
A10 24GB VRAM • | ||||
A100 PCIE 40GB VRAM • Together AI | Not Available | — | ||
A100 PCIE 40GB VRAM • | ||||
A100 SXM 80GB VRAM • SesterceTogether AI | 2x GPU | ↓$1.11(42.7%) | ||
A100 SXM 80GB VRAM • $1.49/hour 2x GPU configuration Updated: 9/8/2026 ★Best Price $2.59/hour Updated: 9/8/2026 Price Difference:↓$1.11(42.7%) | ||||
A16 64GB VRAM • Sesterce | 2x GPU | Not Available | — | |
A16 64GB VRAM • | ||||
A40 48GB VRAM • Sesterce | Not Available | — | ||
A40 48GB VRAM • | ||||
B200 180GB VRAM • Together AI | Not Available | — | ||
B200 180GB VRAM • | ||||
Gaudi 2 96GB VRAM • Sesterce | 8x GPU | Not Available | — | |
Gaudi 2 96GB VRAM • | ||||
H100 SXM 80GB VRAM • SesterceTogether AI | ↓$3.21(59.5%) | |||
H100 SXM 80GB VRAM • $2.19/hour Updated: 9/8/2026 ★Best Price $5.40/hour Updated: 9/8/2026 Price Difference:↓$3.21(59.5%) | ||||
H200 141GB VRAM • SesterceTogether AI | 8x GPU | ↓$2.20(33.3%) | ||
H200 141GB VRAM • $4.40/hour 8x GPU configuration Updated: 9/8/2026 ★Best Price $6.60/hour Updated: 9/8/2026 Price Difference:↓$2.20(33.3%) | ||||
L4 24GB VRAM • Sesterce | Not Available | — | ||
L4 24GB VRAM • | ||||
L40 40GB VRAM • SesterceTogether AI | ↓$0.52(34.9%) | |||
L40 40GB VRAM • $0.97/hour Updated: 9/8/2026 ★Best Price $1.49/hour Updated: 9/8/2026 Price Difference:↓$0.52(34.9%) | ||||
L40S 48GB VRAM • Together AI | Not Available | — | ||
L40S 48GB VRAM • | ||||
RTX 4090 24GB VRAM • Sesterce | Not Available | — | ||
RTX 4090 24GB VRAM • | ||||
RTX 5090 32GB VRAM • Sesterce | Not Available | — | ||
RTX 5090 32GB VRAM • | ||||
RTX 6000 Ada 48GB VRAM • SesterceTogether AI | ↓$0.42(28.3%) | |||
RTX 6000 Ada 48GB VRAM • $1.07/hour Updated: 9/8/2026 ★Best Price $1.49/hour Updated: 9/8/2026 Price Difference:↓$0.42(28.3%) | ||||
A10 24GB VRAM • Sesterce | Not Available | — | ||
A10 24GB VRAM • | ||||
A100 PCIE 40GB VRAM • Together AI | Not Available | — | ||
A100 PCIE 40GB VRAM • | ||||
A100 SXM 80GB VRAM • SesterceTogether AI | 2x GPU | ↓$1.11(42.7%) | ||
A100 SXM 80GB VRAM • $1.49/hour 2x GPU configuration Updated: 9/8/2026 ★Best Price $2.59/hour Updated: 9/8/2026 Price Difference:↓$1.11(42.7%) | ||||
A16 64GB VRAM • Sesterce | 2x GPU | Not Available | — | |
A16 64GB VRAM • | ||||
A40 48GB VRAM • Sesterce | Not Available | — | ||
A40 48GB VRAM • | ||||
B200 180GB VRAM • Together AI | Not Available | — | ||
B200 180GB VRAM • | ||||
Gaudi 2 96GB VRAM • Sesterce | 8x GPU | Not Available | — | |
Gaudi 2 96GB VRAM • | ||||
H100 SXM 80GB VRAM • SesterceTogether AI | ↓$3.21(59.5%) | |||
H100 SXM 80GB VRAM • $2.19/hour Updated: 9/8/2026 ★Best Price $5.40/hour Updated: 9/8/2026 Price Difference:↓$3.21(59.5%) | ||||
H200 141GB VRAM • SesterceTogether AI | 8x GPU | ↓$2.20(33.3%) | ||
H200 141GB VRAM • $4.40/hour 8x GPU configuration Updated: 9/8/2026 ★Best Price $6.60/hour Updated: 9/8/2026 Price Difference:↓$2.20(33.3%) | ||||
L4 24GB VRAM • Sesterce | Not Available | — | ||
L4 24GB VRAM • | ||||
L40 40GB VRAM • SesterceTogether AI | ↓$0.52(34.9%) | |||
L40 40GB VRAM • $0.97/hour Updated: 9/8/2026 ★Best Price $1.49/hour Updated: 9/8/2026 Price Difference:↓$0.52(34.9%) | ||||
L40S 48GB VRAM • Together AI | Not Available | — | ||
L40S 48GB VRAM • | ||||
RTX 4090 24GB VRAM • Sesterce | Not Available | — | ||
RTX 4090 24GB VRAM • | ||||
RTX 5090 32GB VRAM • Sesterce | Not Available | — | ||
RTX 5090 32GB VRAM • | ||||
RTX 6000 Ada 48GB VRAM • SesterceTogether AI | ↓$0.42(28.3%) | |||
RTX 6000 Ada 48GB VRAM • $1.07/hour Updated: 9/8/2026 ★Best Price $1.49/hour Updated: 9/8/2026 Price Difference:↓$0.42(28.3%) | ||||
| Model ↑ | Sesterce | Together AI | Input Diff ↕ |
|---|---|---|---|
DeepSeek | Not available | $0.800 in $0.800 out | — |
DeepSeek | Not available | $3.00 in $7.00 out | — |
DeepSeek | Not available | $2.00 in $2.00 out | — |
DeepSeek | Not available | $0.180 in $0.180 out | — |
DeepSeek | Not available | $1.60 in $1.60 out | — |
DeepSeek | Not available | $0.600 in $1.70 out | — |
DeepSeek | Not available | $0.140 in $0.280 out | — |
DeepSeek | Not available | $1.74 in $3.48 out | — |
Google | Not available | $0.800 in $0.800 out | — |
Google | Not available | $0.390 in $0.970 out | — |
Zhipu | Not available | $0.200 in $1.10 out | — |
Zhipu | Not available | $0.600 in $2.20 out | — |
Zhipu | Not available | $0.450 in $2.00 out | — |
Zhipu | Not available | $1.00 in $3.20 out | — |
Zhipu | Not available | $1.40 in $4.40 out | — |
Explore how these providers compare to other popular GPU cloud services
Compare Sesterce with another leading provider
Compare Sesterce with another leading provider
Compare Sesterce with another leading provider
Compare Sesterce with another leading provider
Compare Sesterce with another leading provider
Compare Sesterce with another leading provider
Each site is operated as a single system, owning the hand-off between energy, facility, silicon, network fabric, and orchestration
French sites offering data residency, controlled energy origin, and supply-chain control, powered by hydro, nuclear, and solar under long-term PPAs
Single console and API to provision clusters, orchestrate training, serve inference, and meter usage across every Sesterce site, Slurm and Kubernetes native
Purpose-built Tier III+ facilities with direct-to-chip liquid cooling at up to 150 kW per rack
Non-blocking NVIDIA Quantum InfiniBand at 800 Gb/s with rail-optimized topology and native RDMA
Dedicated clusters go from signed contract to live hardware in under 2 hours, with single-tenant hardware, network, and storage
Access to Llama, DeepSeek, Qwen, and other leading open-source models
Pay-per-token API with OpenAI-compatible endpoints
LoRA and full fine-tuning with proprietary optimizations
Instant self-service or reserved dedicated clusters with H100, H200, B200, GB200, GB300 access
50% cost reduction for non-urgent inference workloads
Reserve dedicated capacity in throughput units (PTUs) with SLAs
On-demand VMs and bare-metal servers launched with no commitment
Dedicated multi-node GPU clusters ranging from around 100 to 15,000 GPUs
Managed inference instances for serving models alongside GPU compute
Pay-per-second usage with no commitment on VMs and bare-metal servers
Discounted interruptible capacity, filterable as a spot offer in the compute catalog
Committed multi-megawatt cluster capacity contracted for frontier-scale workloads
Pre-funded account balance system for streamlined resource management
Per-token pricing scales based on model size, from small open-source models to 405B parameter frontier models
50% discount for non-urgent inference workloads
Reserve dedicated capacity priced in throughput units (PTUs) with guaranteed SLAs
Per-token pricing for LoRA and full fine-tuning based on model size and dataset
Hourly GPU pricing for instant self-service clusters
Custom pricing for reserved capacity with significant discounts for longer commitments
Single-tenant GPU instances with guaranteed performance
Sign up for Sesterce Cloud platform access
Fund your account with credits for pay-as-you-go billing
Choose your GPU type and configuration, launch in just a few clicks or via the CLI and API
Sign up at together.ai
Generate an API key from your dashboard
Browse 100+ models for chat, code, images, video, and audio
Use OpenAI-compatible endpoints or Together SDK
French footprint: AI factory sites at Bosquel (Hauts-de-France) and Valence, plus a Paris cloud region hosted at Equinix; several on-demand GPU types are offered in more than one region
Documentation and API reference, self-service console with help and requests, REST API and CLI, contact sales for reserved capacity
Global data center network across 25+ cities with frontier hardware including GB300, GB200, B200, H200, H100
Documentation, community Discord, email support, and expert support for reserved cluster customers