On-wafer fabric per engine, 214 Pb/s; cluster networking via dedicated NET cabinets
A thirty-two-node Cerebras inference cluster, provisioned as twenty cabinets: sixteen WSE cabinets holding two CS-3 systems each, plus four NET cabinets. Provisioned heat load is approximately 945 kW, about 875 kW on water and 70 kW on air, Cerebras's own provisioned total including networking. The cluster holds 1.408 TB of on-chip SRAM and no DRAM. Each engine is rated at 125 petaFLOPS, footnoted by Cerebras as sparse; the sparsity is unstructured, so no dense figure is derived.