Code & Dev Tools · verified 2026-07-27
Groq
Absurdly fast inference on custom silicon.
LPU-based serving that returns hundreds of tokens per second on open models, with a free developer tier and an OpenAI-compatible API.
OPEN GROQ- Rating
- 4.6 / 5
- Pricing
- Freemium
- Best for
- Low latency, Open models, Developers
- Platforms
- Web, API
- Runs locally
- No
- Content policy
- Standard filters
Matching filters
What stands out
- Extremely high tokens per second
- Free developer tier
- Drop-in OpenAI API