Claude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/MClaude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/M

Cerebras

Builder of the wafer-scale engine — the largest chip ever made — for fast AI training and inference.

Founded
2015
Headquarters
Sunnyvale, California, US

Cerebras Systems designs the Wafer-Scale Engine, a processor the size of an entire silicon wafer and the largest chip ever built, aimed at both AI training and ultra-fast inference. Its systems avoid the multi-chip communication overhead of GPU clusters, and its inference cloud posts some of the fastest tokens-per-second figures in the industry for large open-weight models.

Founded in 2015 by Andrew Feldman and colleagues (who had previously sold SeaMicro to AMD), Cerebras bet that a single enormous chip could outperform networks of smaller ones for AI. Its Wafer-Scale Engine packs hundreds of thousands of cores and enormous on-chip memory bandwidth onto one wafer-sized die.

The CS-series systems and Cerebras's inference cloud target workloads where speed matters most, delivering very high throughput on large models and attracting customers in research, healthcare, and sovereign-AI programs. Deep partnerships in the Middle East, particularly with the UAE's G42, have been central to its growth and compute deployments.

Cerebras is one of the highest-profile AI-hardware challengers to NVIDIA, competing with Groq and others on the argument that specialized architectures win on inference economics. Its wafer-scale approach remains the most physically distinctive bet in AI silicon.

Cerebras sells model access we track in the pricing index. API pricing & models →

Key people

1

Related companies

3

Part of the Tokenando AI Landscape · Explore the landscape → · All companies →