Developer plan is free for solo users with up to 5k base traces per month, then pay-as-you-go. Plus plan costs $39/seat/month with 10k base traces included. Enterprise plan offers custom pricing with advanced options.
Key features
GroqMetal: dedicated bare-metal infrastructure tuned for speed and reliability
GroqCore: production-ready inference stack with no infrastructure expertise required
GroqAssured: enterprise-grade governance, auditability, and control
256 LPUs per rack with 40 PB/s SRAM bandwidth
1,000 tokens per second per user throughput
128 GB on-chip SRAM per rack
315 PFLOPS of FP8 inference compute
13 globally distributed data centers across four continents
Agent observability with step-by-step tracing
Agent evaluation and scoring
Agent deployment infrastructure
Production monitoring with online evaluations
Autonomous agent improvement with Engine
Secure sandbox execution for agent-generated code
LLM Gateway for model call control
No-code agent builder Fleet
Open source frameworks for agent development
What makes it different
Fast inference at scale without sacrificing affordability
LPU pioneered technology integrated with NVIDIA GPUs
Fully integrated platform combining infrastructure, inference, and control
Comprehensive platform covering the entire agent development lifecycle from building to monitoring
Open source frameworks with options from quick-start to low-level control
Built-in observability and evaluation capabilities integrated into the platform
Support for long-running autonomous agents with Deep Agents framework