NVIDIA Quadro RTX 5000 16GB vs NVIDIA GeForce RTX 2080 Super 8GB

The Quadro RTX 5000 delivers 2.0x the FP16 throughput of the GeForce RTX 2080 Super (89 vs 45 TFLOPS dense). The Quadro RTX 5000 also carries 8 GB more memory (16 GB vs 8 GB).

Quadro RTX 5000: Turing, 2018 GeForce RTX 2080 Super: Turing, 2019
FP16 dense lead
2.0x
Quadro RTX 5000: 89 vs 45 TFLOPS
VRAM
16 vs 8 GB
8 GB more for the Quadro RTX 5000
Bandwidth
448 GB/s vs 496 GB/s
GeForce RTX 2080 Super moves data faster
TDP
265 W vs 250 W
GeForce RTX 2080 Super draws 15 W less

Specifications

Quadro RTX 5000
GeForce RTX 2080 Super
Architecture
TuringTuring
Launch Year
20182019
Form Factor
PCIePCIe
VRAM
16 GB8 GB
Memory Bandwidth
448 GB/s496 GB/s
TDP
265 W250 W
Process Node
12nm12nm

Performance (TFLOPS)

Quadro RTX 5000
GeForce RTX 2080 Super
FP64
0.3 TFlops 0.3 TFlops
FP32
11.2 TFlops 11.2 TFlops
TF32
No verified data available No verified data available
BF16
No verified data available No verified data available
FP16
89.2 TFlops 44.6 TFlops
FP8
No verified data available No verified data available
FP6
No verified data available No verified data available
FP4
No verified data available No verified data available
INT8
178.4 TFlops 178.4 TFlops

FLOPS by Precision

What actually differs

The GeForce RTX 2080 Super is the newer part: Turing, launched in 2019, against the Quadro RTX 5000's Turing from 2018. Newer architectures typically add lower-precision formats and better throughput per watt, so check the precision rows your workload actually uses.

Dense throughput for the Quadro RTX 5000 against the GeForce RTX 2080 Super: FP64 0.3 vs 0.3 TFLOPS, FP32 11 vs 11 TFLOPS, FP16 89 vs 45 TFLOPS. Sparse figures, where the vendor publishes them, appear under each dense number in the performance table.

Memory is 16 GB against 8 GB, fed at 448 GB/s versus 496 GB/s. More memory per GPU means larger models fit before you have to shard across cards, which often matters more than raw TFLOPS for inference.

Power budgets are 265 W for the Quadro RTX 5000 and 250 W for the GeForce RTX 2080 Super. At FP32 that works out to 0.04 against 0.04 TFLOPS per watt.

For LLM training and inference, weight the FP8 and FP16 rows and memory capacity most heavily. For scientific and HPC workloads, the FP64 row is the one to read.

Frequently asked questions

Is the Quadro RTX 5000 faster than the GeForce RTX 2080 Super?

At FP16 precision the Quadro RTX 5000 reaches 89 TFLOPS dense against 45 TFLOPS for the GeForce RTX 2080 Super. The performance table on this page lists every published precision for both GPUs.

Which has more memory, the Quadro RTX 5000 or the GeForce RTX 2080 Super?

The Quadro RTX 5000 carries 16 GB of VRAM versus 8 GB for the GeForce RTX 2080 Super. Memory bandwidth is 448 GB/s for the Quadro RTX 5000 and 496 GB/s for the GeForce RTX 2080 Super.

How much power do the Quadro RTX 5000 and the GeForce RTX 2080 Super draw?

The Quadro RTX 5000 is rated at 265 W TDP and the GeForce RTX 2080 Super at 250 W. On FP32 throughput per watt, the GeForce RTX 2080 Super is the more efficient part.

Get Comparison Updates

New GPUs added weekly. Be the first to see how they compare.