On-wafer fabric per engine, 214 Pb/s; cluster networking via dedicated NET cabinets
The largest Cerebras inference cluster configuration published, at eighty-eight nodes provisioned as fifty-two cabinets: forty-four WSE cabinets holding two CS-3 systems each, plus eight NET cabinets. Provisioned heat load is approximately 2.54 MW, about 2.4 MW on water and 0.14 MW on air, Cerebras's own provisioned total including networking. Fully loaded a WSE cabinet weighs 843 kg and a NET cabinet 963 kg. The cluster holds 3.872 TB of on-chip SRAM and no DRAM at all. Each engine is rated at 125 petaFLOPS, footnoted by Cerebras as sparse; because that sparsity is unstructured rather than 2:4 structured, no dense equivalent can be derived and none is published.