Groq
Groq
An inference cloud built on its own LPU processor, operating dedicated capacity at scale.

Platforms
Overview
Groq calls itself "the premier neocloud for fast inference." It is built around its own LPU processor; the next-generation LPX pairs 256 LPU accelerators per rack with NVIDIA's next-generation GPUs for low-latency, large-context inference. It is sold in three layers, each including the ones below: GroqMetal for bare-metal infrastructure, GroqCore for the inference stack, and GroqAssured for enterprise governance and auditability. The site states it operates 13 data centers across four continents. Pricing is not published.
Main features
- Fast inference on its own LPU processor
- LPX paired with next-generation NVIDIA GPUs
- GroqMetal bare-metal infrastructure
- GroqCore inference stack
- GroqAssured enterprise controls
- 13 data centers across four continents
Pricing
Contact sales
Contact sales
- GroqMetal / GroqCore / GroqAssured
- No published rate card
- Contact for a quote
No price has been verified for your country yet. Check the official site for local availability and pricing.
Screenshots

Related AI tools
Together AI
Together AI is an AI infrastructure and API service useful for AI development and inference.
Fireworks AI
Inference, fine-tuning and dedicated deployments for open models, billed by tokens and GPU hours.
Replicate
Replicate is an AI infrastructure and API service useful for AI development and prototyping.
OpenRouter
OpenRouter is an AI infrastructure and API service useful for AI development and model comparison.