Starter plan is free with $30/month compute credits. Team plan is $250/month plus compute. Enterprise plan is custom. Compute is billed per second for CPU, GPU, and memory with no charges for idle time.
Not given
Key features
Serve LLMs, image, video, and audio models
Isolated environments for coding agents and RL rollouts
Fine-tune and train models on GPUs
Autoscale compute from zero to thousands of GPUs
Distributed storage for models and weights
Global capacity across 20+ clouds
Serverless functions with HTTPS endpoints
Out-of-the-box observability and logs
Find files across Finder, Gmail, Slack, and Google Calendar
Send emails and schedule meetings with voice commands
Review actions before execution or enable auto-execute
Handle multi-step workflows with a single voice request
Works across 20+ apps including Slack, Gmail, Cursor, and Notion
Automatic grammar correction and filler word removal
What makes it different
Custom infrastructure built for AI including container runtime, storage, and scheduler
Sub-second container startup times for GPU workloads
Pay only by the second with no idle charges
Programmatically create millions of concurrent sandboxes
Turn voice into action across multiple apps without switching
Execute multi-step workflows with a single voice command
User maintains full control with review-before-running option
Integrates seamlessly with existing productivity tools and workflows