# Groq
> The fastest LLM inference on the planet (LPU hardware)
- Website: https://groq.com
- Pricing: Free tier; per-token, very cheap
Groq serves open models (Llama, Qwen) at hundreds of tokens/sec on custom LPU chips — for AI features where response speed IS the feature.