NVIDIA HGX Rubin NVL8
Eight Rubin SXM on an HGX baseboard, listed by NVIDIA alongside HGX B300 and HGX B200. NVIDIA publishes NVFP4 twice and the two are different workloads, not a sparse/dense pair: inference 400 PFLOPS sparse, training 280 PFLOPS dense. FP8/FP6 training is 140 PFLOPS dense. System power is not published for the baseboard; the 24 kW figure on DGX Rubin NVL8 is the full chassis including CPUs. NVIDIA marks the specification preliminary.
We will point you at suppliers who have it. Free, and no signup.
Systems built on this design
1 configurationNVIDIA HGX Rubin NVL8 is the baseboard. The systems above are OEM implementations of it and share its accelerator performance; they differ in chassis, CPU, memory and cooling.
System Details
GPU Configuration
System Specifications
Precision Performance Breakdown
| Precision | System Performance | Per GPU | Efficiency |
|---|---|---|---|
| 140,000 TFLOPS | 17,500 TFLOPS | — | |
| 280,000 TFLOPS | 35,000 TFLOPS | — | |
| FP6 | 140,000 TFLOPS | 17,500 TFLOPS | — |
Where a vendor states a basis, the figure shown is the dense one, and any figure it publishes is listed separately. Vendors commonly headline the sparse number instead: for NVIDIA's 2:4 structured sparsity that is exactly twice the dense figure, though other vendors' sparse modes do not all follow that ratio.
Powered by NVIDIA Rubin
This system utilizes 8 × NVIDIA Rubin SXM GPUs, each delivering exceptional performance for AI and HPC workloads.
Per GPU TDP
—
Per GPU Memory
288 GB
Process Node
—
Architecture
Rubin
Documentation & Resources
Flopper Spec Sheet
NVIDIA HGX Rubin NVL8 specifications, generated from our database. Printable.
Vendor product page
NVIDIA HGX Rubin NVL8 documentation
NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition Datasheet
NVIDIA • 2026-06-01
NVIDIA DGX Documentation
System guides, deployment resources
Typical Use Cases
The NVIDIA HGX Rubin NVL8 runs 8× Rubin GPUs over NVLink 6 Switch, 28.8 TB/s total, delivering 140,000 TFLOPS FP8 dense.