Cerebras logo

Cerebras CS-3 Inference Cluster (64-node)

cluster
64× WSE-3
On-wafer fabric per engine, 214 Pb/s; cluster networking via dedicated NET cabinets
2024
Power
1860.0
kW Total
Memory
—
GB Total

A sixty-four-node Cerebras inference cluster, provisioned as thirty-eight cabinets: thirty-two WSE cabinets holding two CS-3 systems each, plus six NET cabinets. Provisioned heat load is approximately 1.86 MW, about 1.75 MW on water and 0.11 MW on air, Cerebras's own provisioned total including networking. The cluster holds 2.816 TB of on-chip SRAM and no DRAM. Each engine is rated at 125 petaFLOPS, footnoted by Cerebras as sparse; the sparsity is unstructured, so no dense figure is derived.

We will point you at suppliers who have it. Free, and no signup.

System Details

GPU Configuration

GPU Model: Cerebras WSE-3
GPU Count: 64 GPUs
Architecture: Wafer Scale Engine 3
Interconnect: On-wafer fabric per engine, 214 Pb/s; cluster networking via dedicated NET cabinets

System Specifications

Form Factor: cluster
Total Power: 1860.0 kW
Total Memory: —
On-die SRAM: 2.82 TB
Memory Bandwidth: 1344 PB/s

Powered by Cerebras WSE-3

This system utilizes 64 × Cerebras WSE-3 GPUs, each delivering exceptional performance for AI and HPC workloads.

Per GPU TDP

—

Per GPU Memory

—

Process Node

TSMC 5nm

Architecture

Wafer Scale Engine 3

Documentation & Resources

Flopper Spec Sheet

Cerebras CS-3 Inference Cluster (64-node) specifications, generated from our database. Printable.

Vendor product page

Cerebras CS-3 Inference Cluster (64-node) documentation

View ↗

Cerebras CS-4 datasheet

Cerebras • 2026-08-18

View Document ↗

Typical Use Cases

AI/ML Training
High-Performance Computing
Data Analytics

The Cerebras CS-3 Inference Cluster (64-node) runs 64× WSE-3 GPUs over On-wafer fabric per engine, 214 Pb/s; cluster networking via dedicated NET cabinets.

© 2026 Flopper.io - Compare the hardware powering AI