Providers /Cerebras
Inference provider · HQ US

Cerebras

Models
3
Regions
00

Our take

Editorial by LLMap · updated Aug 2, 2026

Cerebras is another custom-chip inference provider, using custom chips rather than standard graphics chips. It has an even smaller catalogue than Groq and is aimed at maximum single-stream speed on a handful of supported models.

Use this for maximum single-stream speed on one of its three supported models, or real-time applications where generation speed is the product feature. Skip it if you need a broad catalogue, competitive pricing, or verified EU data residency.

Strengths
  • Purpose-built wafer-scale inference hardware designed for very high tokens per second.
Trade-offs
  • Only three tracked offers, a very small catalogue.
  • Premium pricing on its models compared with other hosts for comparable downloadable models.
01

Privacy & data handling

Yes
No
No answer on record

The colour says whether the answer favours you, not whether it is a yes — not training on your prompts earns a green cross, no zero-retention option earns a red one.

Trains on your prompts
No
Logs prompts
No
Offered
Attested

Prompt retention: none.

02

Models & pricing

Something wrong on this page? Tell us