Today marks the launch of the fourth generation of our Cerebras System, known as CS-4. Recognized as the industry's fastest AI accelerator, this system sets the stage for advanced AI capabilities. It features three newly developed Wafer Scale Engine 3 Turbo processors, integrating a more robust processing unit with a completely reimagined rack and system architecture. CS-4 boasts up to 30 times the inference speed of traditional GPU setups, providing superior cost-effectiveness and a streamlined pathway for deploying hyperscale capacity.
The design philosophy behind CS-4 is grounded in a fundamental principle: significant advancements in AI infrastructure cannot arise from enhancing a single component in isolation. Instead, compute, power, cooling, and I/O must advance in tandem. This holistic design empowers the system to generate rapid tokens, facilitating highly interactive experiences while also ensuring the extensive token capacity required by large-scale operators.
This synergy is crucial across the AI sector:
Developers are eager for responsive reasoning capabilities and dynamic applications at 30 times the current speeds. Data center operators are in search of increased throughput per gigawatt (GW). Furthermore, neoclouds and hyperscalers require adaptable systems that can be efficiently manufactured, installed, expanded, and upgraded on a gigawatt level.
CS-4 embodies these priorities within a single rack-scale platform.



