The Enterprise Agent Build & Runtime for the work your business runs on
Stage
Established
Not given
Founded
Not given
Not given
Based in
US
Not given
Pricing model
Not given
freemium
Pricing
Not given
Free plan with 50 workflow executions per month and visual editor. Enterprise plan with custom pricing includes governance, dedicated VPC, and comprehensive support.
Key features
GroqMetal: dedicated bare-metal infrastructure tuned for speed and reliability
GroqCore: production-ready inference stack with no infrastructure expertise required
GroqAssured: enterprise-grade governance, auditability, and control
256 LPUs per rack with 40 PB/s SRAM bandwidth
1,000 tokens per second per user throughput
128 GB on-chip SRAM per rack
315 PFLOPS of FP8 inference compute
13 globally distributed data centers across four continents
Visual editor and AI copilot
No-code visual editor, exportable to Python
Code-first API built for total control
Real-time tracing of every LLM call, tool call, and memory read
RBAC and audit with immutable audit trails and Enterprise IAM
Human-in-the-loop approval gates and intervention during execution
Runtime hooks inject PII redaction and policy checks
Automated and human-guided training for continuous improvement
Multi-LLM testing for model swapping at runtime
GitHub integration
What makes it different
Fast inference at scale without sacrificing affordability
LPU pioneered technology integrated with NVIDIA GPUs
Fully integrated platform combining infrastructure, inference, and control
Agentic use case generator powered by billions of agent runs
Intelligently guided by 700k agent workflow patterns
Control Plane sits in execution path ensuring every agent interaction is observable, compliant, and reversible
Every production run turns into training data to sharpen accuracy and save money