The Fastest AI Just Got Faster. Introducing the all new Cerebras CS-4, a revolutionary rack-scale solution that delivers up to 30x faster inference compared to GPUs, enhanced economics, and a simple path to deploy hyperscale capacity. It is the architecture for frontier AI. Three WSE-3 Turbo per System Each wafer delivers up to 2x the speed of the previous generation More Performance per Wafer All new power, cooling, and I/O unleashes even more performance per wafer Nexus Rack-Scale Platform Enables rapid deployment in hyperscale datacenters Up to 30x faster than GPUs Powered by WSE-Turbo, CS-4 delivers up to 30x faster inference compared to GPU systems, setting a new record for the fastest inference available in production. Higher ultrafast throughput The CS-4 solution shifts the inference Pareto frontier, delivering up to 10x more throughput per watt than CS-3 while generating tokens up to 30x faster than production GPU systems.…