Cerebras Inference

Wafer-scale inference for very high throughput.

Visit site
Proprietary usage
LLM provider
In these stacks
Write code with AIBuild my own agent
LicenseProprietary
PriceFree tier
UsageMetered
Model supportmulti
Languages
GitHub stars
Last activity
Verified2026-06-27
Verified byseed

An inference provider using wafer-scale CS-3 systems to serve models with very high token throughput, offering a free tier.

Strengths

  • Extremely high token throughput
  • Free tier for experimentation

Tradeoffs

  • Limited to supported models
  • Newer production track record