Today, we are introducing the fourth generation of our Cerebras System: CS-4. The fastest AI accelerator in the industry, and a new foundation for frontier AI. Built from three new Wafer Scale Engine 3 Turbo processors, CS-4 pairs a more powerful processor with a completely redesigned new rack and system. It delivers up to 30 times faster inference than GPU systems, better economics, and a simple path to deploy hyperscale capacity.

CS-4 is designed around a simple idea: the next leap in AI infrastructure cannot come from improving one component in isolation. Compute, power, cooling, and I/O have to move forward together. The result is a system built to generate fast tokens for highly interactive experiences, while also delivering the total token capacity that large-scale operators need.

That combination matters across the AI landscape:

Developers want responsive reasoning and agentic applications at 30x speed.Data center operators want higher throughput per gigawatt (GW.) Neoclouds and hyperscalers need modular systems that can be quickly manufactured, installed, expanded, and upgraded at gigawatt scale.

CS-4 brings those priorities to life in one rack-scale platform.