Serverless inference with per-token pricing (pay for output, input, and cached tokens). Training pricing from $0.50 to $10.00 per 1M training tokens depending on model size and fine-tuning type. On-demand deployments at $0.134 to $0.334 per minute ($8 to $20 per hour) for various GPU types.
Not given
Key features
Guided training runs
Configuration-led training
Custom training logic
Serverless inference
On-demand dedicated deployments
Reserved capacity
OpenAI compatible API
Model library with latest open models
Reinforcement learning support
Multi-region deployments
Build apps and dashboards from prompts without coding
Integrates with Slack, Google Drive, Gmail, Notion, and custom data via MCP and OAuth
Maintains context and picks up where you left off
Runs tasks on a schedule automatically
Search and analyze data from PDFs, Google Drive, and Sheets in plain language
Built on enterprise-grade infrastructure with GDPR compliance
No onboarding or configuration required
Send scheduled Slack alerts and digests
What makes it different
Drop-in replacement for closed-model APIs with cost savings of 50-75%
Instant production deployment from any training checkpoint in seconds
Optimized inference engine for industry-leading throughput and latency
Elastic and global RL inference scaling
Own the complete learning loop for specialized intelligence
No setup or configuration needed to start using
Remembers context and continues work without recaps
Learns tasks and runs them automatically on a schedule
Built by Prosus, one of the world's largest technology groups