NVIDIA GeForce RTX 2080 8GB vs NVIDIA Quadro RTX 6000 24GB

The Quadro RTX 6000 delivers 3.1x the FP16 throughput of the GeForce RTX 2080 (131 vs 42 TFLOPS dense). The Quadro RTX 6000 also carries 16 GB more memory (24 GB vs 8 GB).

GeForce RTX 2080: Turing, 2018 Quadro RTX 6000: Turing, 2018
FP16 dense lead
3.1x
Quadro RTX 6000: 131 vs 42 TFLOPS
VRAM
8 vs 24 GB
16 GB more for the Quadro RTX 6000
Bandwidth
448 GB/s vs 672 GB/s
Quadro RTX 6000 moves data faster
TDP
225 W vs 295 W
GeForce RTX 2080 draws 70 W less

Specifications

GeForce RTX 2080
Quadro RTX 6000
Architecture
TuringTuring
Launch Year
20182018
Form Factor
PCIePCIe
VRAM
8 GB24 GB
Memory Bandwidth
448 GB/s672 GB/s
TDP
225 W295 W
Process Node
12nm12nm

Performance (TFLOPS)

GeForce RTX 2080
Quadro RTX 6000
FP64
0.3 TFlops 0.5 TFlops
FP32
10.6 TFlops 16.3 TFlops
TF32
No verified data available No verified data available
BF16
No verified data available No verified data available
FP16
42.4 TFlops 130.5 TFlops
FP8
No verified data available No verified data available
FP6
No verified data available No verified data available
FP4
No verified data available No verified data available
INT8
169.6 TFlops 261.0 TFlops

FLOPS by Precision

What actually differs

Both GPUs launched in 2018: the GeForce RTX 2080 on Turing and the Quadro RTX 6000 on Turing.

Dense throughput for the GeForce RTX 2080 against the Quadro RTX 6000: FP64 0.3 vs 0.5 TFLOPS, FP32 11 vs 16 TFLOPS, FP16 42 vs 131 TFLOPS. Sparse figures, where the vendor publishes them, appear under each dense number in the performance table.

Memory is 8 GB against 24 GB, fed at 448 GB/s versus 672 GB/s. More memory per GPU means larger models fit before you have to shard across cards, which often matters more than raw TFLOPS for inference.

Power budgets are 225 W for the GeForce RTX 2080 and 295 W for the Quadro RTX 6000. At FP32 that works out to 0.05 against 0.06 TFLOPS per watt.

For LLM training and inference, weight the FP8 and FP16 rows and memory capacity most heavily. For scientific and HPC workloads, the FP64 row is the one to read.

Frequently asked questions

Is the GeForce RTX 2080 faster than the Quadro RTX 6000?

At FP16 precision the Quadro RTX 6000 reaches 131 TFLOPS dense against 42 TFLOPS for the GeForce RTX 2080. The performance table on this page lists every published precision for both GPUs.

Which has more memory, the GeForce RTX 2080 or the Quadro RTX 6000?

The Quadro RTX 6000 carries 24 GB of VRAM versus 8 GB for the GeForce RTX 2080. Memory bandwidth is 448 GB/s for the GeForce RTX 2080 and 672 GB/s for the Quadro RTX 6000.

How much power do the GeForce RTX 2080 and the Quadro RTX 6000 draw?

The GeForce RTX 2080 is rated at 225 W TDP and the Quadro RTX 6000 at 295 W. On FP32 throughput per watt, the Quadro RTX 6000 is the more efficient part.

Can I rent the GeForce RTX 2080 or the Quadro RTX 6000 in the cloud?

Yes. Live cloud listings tracked by Flopper start at $0.24 per GPU hour. The rental pricing section on this page lists current providers and rates for both GPUs.

Where to Rent

NVIDIA Quadro RTX 6000 24GB

ProviderConfigurationPrice/GPU-hrChecked
Vast.ai
1× Quadro RTX 6000 24GB Community $0.24 1h ago View →
Lambda Labs
1× Quadro RTX 6000 24GB $0.69 1h ago View →
All Quadro RTX 6000 listings and price history →

Get Comparison Updates

New GPUs added weekly. Be the first to see how they compare.