NVIDIA Quadro RTX 4000 8GB vs NVIDIA GeForce RTX 2080 Ti 11GB

The Quadro RTX 4000 and the GeForce RTX 2080 Ti deliver near-identical FP16 throughput (57 vs 57 TFLOPS dense). The GeForce RTX 2080 Ti also carries 3 GB more memory (11 GB vs 8 GB).

Quadro RTX 4000: Turing, 2018 GeForce RTX 2080 Ti: Turing, 2018
VRAM
8 vs 11 GB
3 GB more for the GeForce RTX 2080 Ti
Bandwidth
416 GB/s vs 616 GB/s
GeForce RTX 2080 Ti moves data faster
TDP
160 W vs 260 W
Quadro RTX 4000 draws 100 W less

Specifications

Quadro RTX 4000
GeForce RTX 2080 Ti
Architecture
TuringTuring
Launch Year
20182018
Form Factor
PCIePCIe
VRAM
8 GB11 GB
Memory Bandwidth
416 GB/s616 GB/s
TDP
160 W260 W
Process Node
12nm12nm

Performance (TFLOPS)

Quadro RTX 4000
GeForce RTX 2080 Ti
FP64
0.2 TFlops 0.4 TFlops
FP32
7.1 TFlops 14.2 TFlops
TF32
No verified data available No verified data available
BF16
No verified data available No verified data available
FP16
57.0 TFlops 56.9 TFlops
FP8
No verified data available No verified data available
FP6
No verified data available No verified data available
FP4
No verified data available No verified data available
INT8
No verified data available227.7 TFlops

FLOPS by Precision

What actually differs

Both GPUs launched in 2018: the Quadro RTX 4000 on Turing and the GeForce RTX 2080 Ti on Turing.

Dense throughput for the Quadro RTX 4000 against the GeForce RTX 2080 Ti: FP64 0.2 vs 0.4 TFLOPS, FP32 7.1 vs 14 TFLOPS, FP16 57 vs 57 TFLOPS. Sparse figures, where the vendor publishes them, appear under each dense number in the performance table.

Memory is 8 GB against 11 GB, fed at 416 GB/s versus 616 GB/s. More memory per GPU means larger models fit before you have to shard across cards, which often matters more than raw TFLOPS for inference.

Power budgets are 160 W for the Quadro RTX 4000 and 260 W for the GeForce RTX 2080 Ti. At FP32 that works out to 0.04 against 0.05 TFLOPS per watt.

For LLM training and inference, weight the FP8 and FP16 rows and memory capacity most heavily. For scientific and HPC workloads, the FP64 row is the one to read.

Frequently asked questions

Is the Quadro RTX 4000 faster than the GeForce RTX 2080 Ti?

They are close on paper: both deliver about 57 TFLOPS of dense FP16 throughput. Memory, bandwidth, and power are the deciding differences.

Which has more memory, the Quadro RTX 4000 or the GeForce RTX 2080 Ti?

The GeForce RTX 2080 Ti carries 11 GB of VRAM versus 8 GB for the Quadro RTX 4000. Memory bandwidth is 416 GB/s for the Quadro RTX 4000 and 616 GB/s for the GeForce RTX 2080 Ti.

How much power do the Quadro RTX 4000 and the GeForce RTX 2080 Ti draw?

The Quadro RTX 4000 is rated at 160 W TDP and the GeForce RTX 2080 Ti at 260 W. On FP32 throughput per watt, the GeForce RTX 2080 Ti is the more efficient part.

Can I rent the Quadro RTX 4000 or the GeForce RTX 2080 Ti in the cloud?

Yes. Live cloud listings tracked by Flopper start at $0.07 per GPU hour. The rental pricing section on this page lists current providers and rates for both GPUs.

Where to Rent

NVIDIA Quadro RTX 4000 8GB

ProviderConfigurationPrice/GPU-hrChecked
Vast.ai
2× Quadro RTX 4000 8GB Community $0.07 4h ago View →
All Quadro RTX 4000 listings and price history →

NVIDIA GeForce RTX 2080 Ti 11GB

ProviderConfigurationPrice/GPU-hrChecked
Vast.ai
8× GeForce RTX 2080 Ti 11GB Community $0.07 4h ago View →
All GeForce RTX 2080 Ti listings and price history →

Get Comparison Updates

New GPUs added weekly. Be the first to see how they compare.