NVIDIA GeForce RTX 5090 32GB vs NVIDIA B100 SXM 192GB

The B100 delivers 4.2x the FP8 throughput of the GeForce RTX 5090 (3,500 vs 838 TFLOPS dense). The B100 also carries 160 GB more memory (192 GB vs 32 GB).

GeForce RTX 5090: Blackwell, 2025 B100: Blackwell, 2024
FP8 dense lead
4.2x
B100: 3,500 vs 838 TFLOPS
Memory
32 vs 192 GB
160 GB more for the B100
Bandwidth
1.8 TB/s vs 8.0 TB/s
B100 moves data faster
TDP
575 W vs 700 W
GeForce RTX 5090 draws 125 W less

Specifications

GeForce RTX 5090
B100
Architecture
BlackwellBlackwell
Launch Year
20252024
Form Factor
PCIeSXM
Memory
32 GB192 GB
Memory Bandwidth
1.8 TB/s8.0 TB/s
TDP
575 W700 W
Process Node
4nm4nm

Performance (TFLOPS)

GeForce RTX 5090
B100
FP64
1.6 TFLOPS 30 TFLOPS
FP32
105 TFLOPS 60 TFLOPS
TF32
105 TFLOPS
210 TFLOPS sparse
875 TFLOPS
sparse not published
BF16
210 TFLOPS
419 TFLOPS sparse
1,750 TFLOPS
sparse not published
FP16
419 TFLOPS
sparse not published
1,750 TFLOPS
sparse not published
FP8
838 TFLOPS
sparse not published
3,500 TFLOPS
sparse not published
FP6
No verified data available No verified data available
FP4
1,676 TFLOPS
sparse not published
7,000 TFLOPS
sparse not published
INT8
838 TOPS
sparse not published
3,500 TOPS
sparse not published

FLOPS by Precision

What actually differs

The GeForce RTX 5090 is the newer part: Blackwell, launched in 2025, against the B100's Blackwell from 2024. Newer architectures typically add lower-precision formats and better throughput per watt, so check the precision rows your workload actually uses.

Dense throughput for the GeForce RTX 5090 against the B100: FP64 1.6 vs 30 TFLOPS, FP32 105 vs 60 TFLOPS, FP16 419 vs 1,750 TFLOPS, FP8 838 vs 3,500 TFLOPS. Sparse figures, where the vendor publishes them, appear under each dense number in the performance table.

Memory is 32 GB against 192 GB, fed at 1.8 TB/s versus 8.0 TB/s. More memory per GPU means larger models fit before you have to shard across cards, which often matters more than raw TFLOPS for inference.

Power budgets are 575 W for the GeForce RTX 5090 and 700 W for the B100. At FP32 that works out to 0.18 against 0.09 TFLOPS per watt.

For LLM training and inference, weight the FP8 and FP16 rows and memory capacity most heavily. For scientific and HPC workloads, the FP64 row is the one to read.

Frequently asked questions

Is the GeForce RTX 5090 faster than the B100?

At FP8 precision the B100 reaches 3,500 TFLOPS dense against 838 TFLOPS for the GeForce RTX 5090. The performance table on this page lists every published precision for both GPUs.

Which has more memory, the GeForce RTX 5090 or the B100?

The B100 carries 192 GB of memory versus 32 GB for the GeForce RTX 5090. Memory bandwidth is 1.8 TB/s for the GeForce RTX 5090 and 8.0 TB/s for the B100.

How much power do the GeForce RTX 5090 and the B100 draw?

The GeForce RTX 5090 is rated at 575 W TDP and the B100 at 700 W. On FP32 throughput per watt, the GeForce RTX 5090 is the more efficient part.

Can I rent the GeForce RTX 5090 or the B100 in the cloud?

Yes. Live cloud listings tracked by Flopper start at $0.35 per GPU hour. The rental pricing section on this page lists current providers and rates for both GPUs.

Where to Rent

NVIDIA GeForce RTX 5090 32GB

ProviderConfigurationPrice/GPU-hrChecked
Vast.ai
1× GeForce RTX 5090 32GB Community $0.35 4h ago View →
RunPod
1× GeForce RTX 5090 32GB Community $0.69 4h ago View →
RunPod
1× GeForce RTX 5090 32GB $0.99 4h ago View →
All GeForce RTX 5090 listings and price history →

Get Comparison Updates

New GPUs added weekly. Be the first to see how they compare.