NVIDIA Vera Rubin Superchip

Rubin Superchip 2026 Spec confidence: Vendor claimed
FP8 (dense)
35,000
TFLOPS
FP32
260
TFLOPS
VRAM
576 GB
HBM4
Bandwidth
44.0 TB/s
memory
Download datasheet

Overview

The Vera Rubin Superchip pairs two Rubin GPUs (576 GB total HBM4, 44 TB/s aggregate memory bandwidth) with one 88-core Vera CPU built on Arm-compatible Olympus cores and 1.5 TB of LPDDR5X CPU memory. The CPU and GPUs are linked by NVLink-C2C at 1.8 TB/s coherent bandwidth, with GPU-to-GPU NVLink bandwidth of 7.2 TB/s aggregate. Headline NVFP4 inference throughput is 100 PFLOPS dense; NVFP4 training is 70 PFLOPS dense. Tensor-core-emulated FP32 (SGEMM) reaches 800 TFLOPS and FP64 (DGEMM) 400 TFLOPS. This Superchip is the rack-building block for Vera Rubin NVL72. Specifications are preliminary per NVIDIA and subject to change.

Performance

Peak theoretical throughput by precision type

FP32_TC
800 TFLOPS
FP6
35,000 TFLOPS
sparse not published
NVFP4
100,000 TFLOPS
sparse not published
INT8
8-bit integer
500 TOPS
sparse not published
FP8
8-bit floating point
35,000 TFLOPS
sparse not published
FP16
16-bit floating point
8,000 TFLOPS
sparse not published
BF16
Brain Float 16
8,000 TFLOPS
sparse not published
TF32
TensorFloat-32
4,000 TFLOPS
sparse not published
FP32
32-bit floating point
260 TFLOPS
FP64_TC
64-bit floating point with Tensor Cores
400 TFLOPS
FP64
64-bit floating point
67 TFLOPS

Specifications

Architecture

Rubin

Form Factor

Superchip

Launch Year

2026

Memory

576 GB HBM4

Bandwidth

44,000 GB/s

Interconnect

1.8 TB/s NVLink-C2C

Spec Confidence

Vendor claimed

Full Specifications

Compute Engine
GPUs per Module 2 (all figures are totals for the module)
Memory
VRAM 576 GB
Memory Type HBM4
Bandwidth 44.0 TB/s
Interconnect & I/O
GPU-to-GPU NVLink-C2C
Interconnect Bandwidth 1.8 TB/s
Power & Thermal
Enterprise Features
Sparsity Yes
General
Form Factor Superchip
Architecture Rubin
Launch Year 2026

Datasheet & Resources

Data Provenance

Every figure traced to a source

Primary Source

Publisher
NVIDIA
Published
2025-03-18

Data Quality

Spec confidence
Vendor claimed
Clock basis
Boost
Precisions tracked
11
Normalization
All values in TFLOPS

Compare cloud providers offering on-demand GPU instances for AI training, inference, and HPC workloads.

Browse GPU Cloud Providers

Similar GPUs

NVIDIA Rubin SXM

SXM · 2026
FP32: 130 TFLOPS
Compare vs Rubin

Frequently Asked Questions

How many TFLOPS does the NVIDIA Vera Rubin have?

The NVIDIA Vera Rubin delivers 260 TFLOPS for FP32 operations, 8000 TFLOPS for FP16, and 35000 TFLOPS for FP8 precision.

What is the power consumption of the NVIDIA Vera Rubin?

The NVIDIA Vera Rubin has a TDP (Thermal Design Power) rating of N/A watts.

How much memory does the NVIDIA Vera Rubin have?

The NVIDIA Vera Rubin is equipped with 576 GB of VRAM with 44000 GB/s memory bandwidth.

What architecture is the NVIDIA Vera Rubin based on?

The NVIDIA Vera Rubin is based on the Rubin architecture, launched in 2026.

Stay Updated on GPU Releases

Get notified when new GPUs are added or specifications are updated.

Loading verification...

No spam, unsubscribe anytime.

Back to GPUs