Cerebras CS-3 Inference Cluster (64-node)
A sixty-four-node Cerebras inference cluster, provisioned as thirty-eight cabinets: thirty-two WSE cabinets holding two CS-3 systems each, plus six NET cabinets. Provisioned heat load is approximately 1.86 MW, about 1.75 MW on water and 0.11 MW on air, Cerebras's own provisioned total including networking. The cluster holds 2.816 TB of on-chip SRAM and no DRAM. Each engine is rated at 125 petaFLOPS, footnoted by Cerebras as sparse; the sparsity is unstructured, so no dense figure is derived.
We will point you at suppliers who have it. Free, and no signup.
System Details
GPU Configuration
System Specifications
Powered by Cerebras WSE-3
This system utilizes 64 × Cerebras WSE-3 GPUs, each delivering exceptional performance for AI and HPC workloads.
Per GPU TDP
—
Per GPU Memory
—
Process Node
TSMC 5nm
Architecture
Wafer Scale Engine 3
Documentation & Resources
Official Datasheet
Cerebras CS-3 Inference Cluster (64-node) technical specifications
Cerebras CS-3 datasheet: system, rack and inference cluster specifications
Cerebras • Latest version
Typical Use Cases
The Cerebras CS-3 Inference Cluster (64-node) runs 64× WSE-3 GPUs over On-wafer fabric per engine, 214 Pb/s; cluster networking via dedicated NET cabinets.