Starter plan is free with $30/month compute credits. Team plan is $250/month plus compute. Enterprise plan is custom. Compute is billed per second for CPU, GPU, and memory with no charges for idle time.
Not given
Key features
Serve LLMs, image, video, and audio models
Isolated environments for coding agents and RL rollouts
Fine-tune and train models on GPUs
Autoscale compute from zero to thousands of GPUs
Distributed storage for models and weights
Global capacity across 20+ clouds
Serverless functions with HTTPS endpoints
Out-of-the-box observability and logs
One user-owned context for every agent
Agents wake up to what changed since their last loop
Readable, traceable context with every change logged
End-to-end encryption with AES-256
Granular sharing controls with revocable access
Right to export, delete, and transparency
Context that travels across AI products and models
What makes it different
Custom infrastructure built for AI including container runtime, storage, and scheduler
Sub-second container startup times for GPU workloads
Pay only by the second with no idle charges
Programmatically create millions of concurrent sandboxes
User-owned context that travels with you across AI products
Traceable agent activity with full visibility and control
Agents coordinate through shared context without learning being trapped in silos