On-demand instances available hourly; 1-Click Clusters from $8.87 to $9.86 per GPU per hour for 2 weeks to 1 year; reserved capacity available at lower prices with annual commitment.
Per-second billing for Pods, Serverless, and Clusters with no minimum commitments. GPU pricing ranges from $0.27/hr to $7.89/hr depending on GPU type and configuration.
Key features
Superclusters with NVIDIA GB300 NVL72 and Quantum-2 InfiniBand
1-Click Clusters™ with NVIDIA HGX B200 and H100 GPUs
On-demand GPU instances
Single-tenant, shared-nothing architecture
SOC 2 Type II certification
Managed cluster orchestration
Co-engineering support
Autoscaling serverless endpoints
Multi-node GPU clusters
Global deployment across 31 regions
High-speed InfiniBand networking
Slurm orchestration support
Docker container support
SOC 2 Type II compliance
What makes it different
Complete AI factories integrating power, cooling, and GPUs in one system designed for peak AI performance
Dedicated single-tenant infrastructure for training and inference at scale
Expert co-engineering from team building infrastructure for world's largest AI labs
Rack-scale systems optimized for agentic AI and reasoning workloads
Hardware-level isolation with caged clusters
No cold start latency or idle costs with production inference without warm-up tax
Instant multi-node cluster deployment in minutes
No contracts or minimum commitments with complete flexibility
Low pricing up to 90% cheaper than hyperscalers
Unified platform for full AI lifecycle from experiment to production