NVIDIA DGX Rubin NVL8

rack
8× Rubin
NVLink 6 Switch, 28.8 TB/s total
2026
FP8
140,000
TFLOPS
NVFP4
280,000
TFLOPS
Power
24.0
kW Total
Memory
2304
GB Total

Eight Rubin GPUs with two Intel Xeon 6776P CPUs, liquid cooled. NVIDIA publishes NVFP4 twice and the two are different workloads, not a sparse/dense pair: inference 400 PFLOPS sparse, training 280 PFLOPS dense. The dense training figure is stored as the headline; the sparse inference figure is recorded alongside it. FP8/FP6 training is 140 PFLOPS dense with no sparse counterpart published, and no INT8 figure is given. Networking is 8x OSFP with ConnectX-9 up to 800 Gb/s plus 2x BlueField-4 DPUs. NVIDIA marks the whole specification "preliminary information, all values subject to change".

We will point you at suppliers who have it. Free, and no signup.

FP8
140,000
TFLOPS
NVFP4
280,000
TFLOPS
FP6
140,000
TFLOPS

System Details

GPU Configuration

GPU Model: NVIDIA Rubin SXM
GPU Count: 8 GPUs
Architecture: Rubin
Interconnect: NVLink 6 Switch, 28.8 TB/s total

System Specifications

Form Factor: rack
Total Power: 24.0 kW
Total Memory: 2.3 TB
Memory Bandwidth: 176 TB/s

Precision Performance Breakdown

PrecisionSystem PerformancePer GPUEfficiency
140,000 TFLOPS 17,500 TFLOPS —
280,000 TFLOPS 35,000 TFLOPS —
FP6 140,000 TFLOPS 17,500 TFLOPS —

Where a vendor states a basis, the figure shown is the dense one, and any figure it publishes is listed separately. Vendors commonly headline the sparse number instead: for NVIDIA's 2:4 structured sparsity that is exactly twice the dense figure, though other vendors' sparse modes do not all follow that ratio.

Powered by NVIDIA Rubin

This system utilizes 8 × NVIDIA Rubin SXM GPUs, each delivering exceptional performance for AI and HPC workloads.

Per GPU TDP

—

Per GPU Memory

288 GB

Process Node

—

Architecture

Rubin

Documentation & Resources

Flopper Spec Sheet

NVIDIA DGX Rubin NVL8 specifications, generated from our database. Printable.

Vendor product page

NVIDIA DGX Rubin NVL8 documentation

View ↗

NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition Datasheet

NVIDIA • 2026-06-01

View Document ↗

NVIDIA DGX Documentation

System guides, deployment resources

Browse ↗

Typical Use Cases

Large Language Model Training
Distributed Deep Learning
Multi-GPU Inference
HPC Simulations
Scientific Computing

The NVIDIA DGX Rubin NVL8 runs 8× Rubin GPUs over NVLink 6 Switch, 28.8 TB/s total, delivering 140,000 TFLOPS FP8 dense.

© 2026 Flopper.io - Compare the hardware powering AI