On-wafer fabric per engine, 214 Pb/s; cluster networking via dedicated NET cabinets
A sixteen-node Cerebras inference cluster, provisioned as ten cabinets: eight WSE cabinets holding two CS-3 systems each, plus two NET cabinets. Provisioned heat load is approximately 475 kW, about 440 kW on water and 35 kW on air. That is Cerebras's own provisioned total including networking, not sixteen times the 27 kW system figure. The cluster holds 704 GB of on-chip SRAM and no DRAM. Each engine is rated at 125 petaFLOPS, footnoted by Cerebras as sparse; the sparsity is unstructured, so no dense figure is derived.