All payments in the preview are in test mode. Read more
← Index

Code & Dev Tools · verified 2026-07-14

Cerebras Inference

Wafer-scale speed for open models.

Inference on Cerebras hardware, competing with Groq on raw throughput for Llama and Qwen-class models.

OPEN CEREBRAS INFERENCE
Rating
4.3 / 5
Pricing
Freemium
Best for
Low latency, Open models, Developers
Platforms
Web, API
Runs locally
No
Content policy
Standard filters

Matching filters

What stands out

  • Very high throughput
  • Free tier for developers
  • OpenAI-compatible endpoint

Also in Code & Dev Tools