NVIDIA GeForce RTX 5090 32GB vs NVIDIA GB200 Grace Blackwell Superchip 372GB

The GB200 delivers 11.9x the FP8 throughput of the GeForce RTX 5090 (10,000 vs 838 TFLOPS dense). The GB200 also carries 340 GB more memory (372 GB vs 32 GB).

GeForce RTX 5090: Blackwell, 2025 GB200: Blackwell, 2024
FP8 dense lead
11.9x
GB200: 10,000 vs 838 TFLOPS
Memory
32 vs 372 GB
340 GB more for the GB200
Bandwidth
1.8 TB/s vs 16.0 TB/s
GB200 moves data faster

Specifications

GeForce RTX 5090
GB200
Architecture
BlackwellBlackwell
Launch Year
20252024
Form Factor
PCIeSuperchip
Memory
32 GB372 GB
Memory Bandwidth
1.8 TB/s16.0 TB/s
TDP
575 W
Process Node
4nm4nm

Performance (TFLOPS)

GeForce RTX 5090
GB200
FP64
1.6 TFLOPS 80 TFLOPS
FP32
105 TFLOPS 160 TFLOPS
TF32
105 TFLOPS
210 TFLOPS sparse
2,500 TFLOPS
5,000 TFLOPS sparse
BF16
210 TFLOPS
419 TFLOPS sparse
5,000 TFLOPS
10,000 TFLOPS sparse
FP16
419 TFLOPS
sparse not published
5,000 TFLOPS
10,000 TFLOPS sparse
FP8
838 TFLOPS
sparse not published
10,000 TFLOPS
20,000 TFLOPS sparse
FP6
No verified data available10,000 TFLOPS
20,000 TFLOPS sparse
FP4
1,676 TFLOPS
sparse not published
20,000 TFLOPS
40,000 TFLOPS sparse
INT8
838 TOPS
sparse not published
10,000 TOPS
20,000 TOPS sparse

FLOPS by Precision

What actually differs

The GeForce RTX 5090 is the newer part: Blackwell, launched in 2025, against the GB200's Blackwell from 2024. Newer architectures typically add lower-precision formats and better throughput per watt, so check the precision rows your workload actually uses.

Dense throughput for the GeForce RTX 5090 against the GB200: FP64 1.6 vs 80 TFLOPS, FP32 105 vs 160 TFLOPS, FP16 419 vs 5,000 TFLOPS, FP8 838 vs 10,000 TFLOPS. Sparse figures, where the vendor publishes them, appear under each dense number in the performance table.

Memory is 32 GB against 372 GB, fed at 1.8 TB/s versus 16.0 TB/s. More memory per GPU means larger models fit before you have to shard across cards, which often matters more than raw TFLOPS for inference.

For LLM training and inference, weight the FP8 and FP16 rows and memory capacity most heavily. For scientific and HPC workloads, the FP64 row is the one to read.

Frequently asked questions

Is the GeForce RTX 5090 faster than the GB200?

At FP8 precision the GB200 reaches 10,000 TFLOPS dense against 838 TFLOPS for the GeForce RTX 5090. The performance table on this page lists every published precision for both GPUs.

Which has more memory, the GeForce RTX 5090 or the GB200?

The GB200 carries 372 GB of memory versus 32 GB for the GeForce RTX 5090. Memory bandwidth is 1.8 TB/s for the GeForce RTX 5090 and 16.0 TB/s for the GB200.

Can I rent the GeForce RTX 5090 or the GB200 in the cloud?

Yes. Live cloud listings tracked by Flopper start at $0.35 per GPU hour. The rental pricing section on this page lists current providers and rates for both GPUs.

Where to Rent

NVIDIA GeForce RTX 5090 32GB

ProviderConfigurationPrice/GPU-hrChecked
Vast.ai
1× GeForce RTX 5090 32GB Community $0.35 4h ago View →
RunPod
1× GeForce RTX 5090 32GB Community $0.69 4h ago View →
RunPod
1× GeForce RTX 5090 32GB $0.99 4h ago View →
All GeForce RTX 5090 listings and price history →

Get Comparison Updates

New GPUs added weekly. Be the first to see how they compare.