CS-4

Overview

Rack-scale AI system on the new Cerebras Nexus platform, built from three WSE-3 Turbo processors in modular liquid-cooled compute backpacks. Cerebras says CS-4 delivers up to 30x faster inference than GPU systems, up to 10x more throughput per watt than CS-3, 750 PFLOPs of AI compute, 7.2 Tbps of I/O, 129.6 PB/s of memory bandwidth, and wafer-to-wafer latency as low as 2 microseconds for models above 50 trillion parameters.

Unveiled August 18, 2026. First customer shipments targeted for Q3 2026. Nexus is also the rack for CS-5, aimed at 2027.

Key Specs

Wafers per rack
3x WSE-3 Turbo
AI compute
750 PFLOPs
Memory bandwidth
129.6 PB/s
Claimed speed vs GPUs
Up to 30x faster inference; up to 10x more throughput per watt than CS-3