Cerebras CS-4 Rack System Runs AI Inference Up to 30 Times Faster Than GPUs
Cerebras unveiled the CS-4 rack-scale solution, claiming it performs AI inference up to 30 times faster than simple GPU systems. The system features three Wafer Scale Engine 3 Turbo processors per system, with each wafer achieving up to twice the speed of the previous generation. Initial shipments begin this quarter.