NVIDIA GeForce RTX 5080 16GB vs NVIDIA B100 SXM 192GB

The B100 delivers 7.8x the FP8 throughput of the GeForce RTX 5080 (3,500 vs 450 TFLOPS dense). The B100 also carries 176 GB more memory (192 GB vs 16 GB).

GeForce RTX 5080: Blackwell, 2025 B100: Blackwell, 2024
FP8 dense lead
7.8x
B100: 3,500 vs 450 TFLOPS
Memory
16 vs 192 GB
176 GB more for the B100
Bandwidth
960 GB/s vs 8.0 TB/s
B100 moves data faster
TDP
360 W vs 700 W
GeForce RTX 5080 draws 340 W less

Specifications

GeForce RTX 5080
B100
Architecture
BlackwellBlackwell
Launch Year
20252024
Form Factor
PCIeSXM
Memory
16 GB192 GB
Memory Bandwidth
960 GB/s8.0 TB/s
TDP
360 W700 W
Process Node
4nm4nm

Performance (TFLOPS)

GeForce RTX 5080
B100
FP64
0.9 TFLOPS 30 TFLOPS
FP32
56 TFLOPS 60 TFLOPS
TF32
56 TFLOPS
113 TFLOPS sparse
875 TFLOPS
sparse not published
BF16
113 TFLOPS
225 TFLOPS sparse
1,750 TFLOPS
sparse not published
FP16
225 TFLOPS
sparse not published
1,750 TFLOPS
sparse not published
FP8
450 TFLOPS
sparse not published
3,500 TFLOPS
sparse not published
FP6
No verified data available No verified data available
FP4
901 TFLOPS
sparse not published
7,000 TFLOPS
sparse not published
INT8
450 TOPS
sparse not published
3,500 TOPS
sparse not published

FLOPS by Precision

What actually differs

The GeForce RTX 5080 is the newer part: Blackwell, launched in 2025, against the B100's Blackwell from 2024. Newer architectures typically add lower-precision formats and better throughput per watt, so check the precision rows your workload actually uses.

Dense throughput for the GeForce RTX 5080 against the B100: FP64 0.9 vs 30 TFLOPS, FP32 56 vs 60 TFLOPS, FP16 225 vs 1,750 TFLOPS, FP8 450 vs 3,500 TFLOPS. Sparse figures, where the vendor publishes them, appear under each dense number in the performance table.

Memory is 16 GB against 192 GB, fed at 960 GB/s versus 8.0 TB/s. More memory per GPU means larger models fit before you have to shard across cards, which often matters more than raw TFLOPS for inference.

Power budgets are 360 W for the GeForce RTX 5080 and 700 W for the B100. At FP32 that works out to 0.16 against 0.09 TFLOPS per watt.

For LLM training and inference, weight the FP8 and FP16 rows and memory capacity most heavily. For scientific and HPC workloads, the FP64 row is the one to read.

Frequently asked questions

Is the GeForce RTX 5080 faster than the B100?

At FP8 precision the B100 reaches 3,500 TFLOPS dense against 450 TFLOPS for the GeForce RTX 5080. The performance table on this page lists every published precision for both GPUs.

Which has more memory, the GeForce RTX 5080 or the B100?

The B100 carries 192 GB of memory versus 16 GB for the GeForce RTX 5080. Memory bandwidth is 960 GB/s for the GeForce RTX 5080 and 8.0 TB/s for the B100.

How much power do the GeForce RTX 5080 and the B100 draw?

The GeForce RTX 5080 is rated at 360 W TDP and the B100 at 700 W. On FP32 throughput per watt, the GeForce RTX 5080 is the more efficient part.

Can I rent the GeForce RTX 5080 or the B100 in the cloud?

Yes. Live cloud listings tracked by Flopper start at $0.18 per GPU hour. The rental pricing section on this page lists current providers and rates for both GPUs.

Where to Rent

NVIDIA GeForce RTX 5080 16GB

ProviderConfigurationPrice/GPU-hrChecked
Vast.ai
1× GeForce RTX 5080 16GB Community $0.18 3h ago View →
RunPod
1× GeForce RTX 5080 16GB Community $0.39 3h ago View →
RunPod
1× GeForce RTX 5080 16GB $0.59 3h ago View →
All GeForce RTX 5080 listings and price history →

Get Comparison Updates

New GPUs added weekly. Be the first to see how they compare.