Cerebras logo

Cerebras

Wafer-scale AI chip company (WSE-3: 2,100 tok/s, 8x faster than H200). 900,000 AI cores on single die. IPO targeting Q2 2026 (~$22B valuation). Powers GPT-5.3-Codex-Spark real-time inference.

Share

Listen to this lesson

Free preview · first 0:30
0:00 / 0:30

Unlock audio and more

Audio streaming, downloadable PDFs and certificates come with Plus and Pro.

📋About Cerebras

Updated September 16, 2026

Cerebras Systems is an AI hardware company founded in 2016, known for building the world's largest computer chips. The company's Wafer-Scale Engine (WSE-3) is a single chip that occupies an entire silicon wafer — containing 4 trillion transistors and 900,000 AI-optimized cores, making it roughly 50 times larger than the largest GPU.

Cerebras's CS-3 systems using the WSE-3 are designed for AI training and inference workloads that benefit from massive on-chip memory and bandwidth, eliminating the communication bottlenecks that occur in multi-GPU clusters. The company also operates Cerebras Inference, a cloud service offering some of the fastest inference speeds available — delivering over 1,800 tokens per second for Llama models, significantly faster than GPU-based alternatives.

Cerebras raised approximately $2.9 billion across 8 rounds privately before going public in May 2026, completing the largest US tech IPO of the year: 28 million shares priced at $185 (above the $115 to $160 range) for roughly $5.5 billion in proceeds, with the stock more than doubling on its first day of trading to close at $311, which valued the company at roughly $66 billion at that closing price. The S-1 named OpenAI, Group 42, Saudi Arabia's MBZUAI (Mohamed bin Zayed University of Artificial Intelligence), and Amazon Web Services as top customers, and Cerebras swung to profitability on $510 million of 2025 revenue. OpenAI is one of Cerebras's largest customers under a multi-year contract worth more than $10 billion signed in January 2026, and holds a $1 billion secured loan plus warrants for over 33 million shares — making OpenAI a meaningful post-listing shareholder. That relationship became publicly visible in the product in August 2026, when OpenAI launched an Ultrafast service tier that serves its GPT-5.6 Sol model from Cerebras wafer-scale hardware rather than GPUs, reaching up to 750 output tokens per second. Cerebras reports a 5.6-times end-to-end speedup on the GDP-Val benchmark with no quality loss. The tier launched in limited preview in the OpenAI API, and it is the clearest demonstration so far of Cerebras serving a frontier lab's flagship model rather than open-weight alternatives. The unconventional approach to AI computing challenges the dominant NVIDIA GPU paradigm, offering an alternative architecture that excels at high-bandwidth answer-inference workloads where on-chip SRAM gives Cerebras a practical latency advantage over HBM-based GPUs. Other major customers include AWS, Meta, Mistral, Perplexity, Mayo Clinic, US defense agencies, and pharmaceutical and research labs that need extreme computational throughput.

Cerebras went public on Nasdaq on May 14, 2026 in the largest US technology listing since Snowflake, pricing 28 million shares at $185 and raising roughly $5.5 billion. On August 18, 2026 it launched the CS-4 rack system, which reaches its speed gains by overclocking the existing wafer-scale engine from 1.4 to 2.8 gigahertz and packing three wafers per rack rather than by fabricating new silicon — a genuine improvement that does not raise the on-wafer memory ceiling that forces frontier models across multiple systems. Alongside it, AMD joined OpenAI as a named partner in a disaggregated inference design that splits serving between vendors, with AMD graphics processors handling the prefill phase and Cerebras handling token generation.

🛠️Products & Tools (1)

Cerebras InferenceInference & Model Serving

AI inference platform powered by wafer-scale processors. The CS-4 rack system (August 2026) runs models above 10 trillion parameters at more than 1,000 tokens per second; the hosted cloud serves open models up to 480 billion parameters. Partners include OpenAI, AMD, AWS and Meta.

Keep track of the companies you’re watching

  • The AI Hub on a phone: a 12-day AI Skill Streak and an expanded Content updates alert listing the saved items that changed.
  • Recommended for you on a phone: nine personalised suggestions labelled Trending in AI news, On your saved list, and Popular.
  • My AI Tools on a phone: saved tools including GitHub Copilot and OpenAI Codex, each with an Updated badge.

Swipe for Recommended for you and My AI Tools

Your AI Hub — sample data.

📰Cerebras in the News

Showing the 5 stories where Cerebras is tagged in Top AI Stories.