NVIDIA B100 SXM 192GB vs NVIDIA GeForce RTX 5080 16GB

The B100 delivers 7.8x the FP8 throughput of the GeForce RTX 5080 (3,500 vs 450 TFLOPS dense). The B100 also carries 176 GB more memory (192 GB vs 16 GB).

B100: Blackwell, 2024 GeForce RTX 5080: Blackwell, 2025
FP8 dense lead
7.8x
B100: 3,500 vs 450 TFLOPS
VRAM
192 vs 16 GB
176 GB more for the B100
Bandwidth
8.0 TB/s vs 960 GB/s
B100 moves data faster
TDP
700 W vs 360 W
GeForce RTX 5080 draws 340 W less

Specifications

B100
GeForce RTX 5080
Architecture
BlackwellBlackwell
Launch Year
20242025
Form Factor
SXMPCIe
VRAM
192 GB16 GB
Memory Bandwidth
8.0 TB/s960 GB/s
TDP
700 W360 W
Process Node
4nm4nm

Performance (TFLOPS)

B100
GeForce RTX 5080
FP64
30.0 TFlops 0.9 TFlops
FP32
60.0 TFlops 56.3 TFlops
FP16
1750.0 TFlops
sparse not published
225.1 TFlops
sparse not published
BF16
1750.0 TFlops
sparse not published
225.1 TFlops
sparse not published
FP8
3500.0 TFlops
sparse not published
450.2 TFlops
sparse not published
INT8
3500.0 TFlops
sparse not published
450.2 TFlops
sparse not published

FLOPS by Precision

What actually differs

The GeForce RTX 5080 is the newer part: Blackwell, launched in 2025, against the B100's Blackwell from 2024. Newer architectures typically add lower-precision formats and better throughput per watt, so check the precision rows your workload actually uses.

Dense throughput for the B100 against the GeForce RTX 5080: FP64 30 vs 0.9 TFLOPS, FP32 60 vs 56 TFLOPS, FP16 1,750 vs 225 TFLOPS, FP8 3,500 vs 450 TFLOPS. Sparse figures, where the vendor publishes them, appear under each dense number in the performance table.

Memory is 192 GB against 16 GB, fed at 8.0 TB/s versus 960 GB/s. More memory per GPU means larger models fit before you have to shard across cards, which often matters more than raw TFLOPS for inference.

Power budgets are 700 W for the B100 and 360 W for the GeForce RTX 5080. At FP32 that works out to 0.09 against 0.16 TFLOPS per watt.

For LLM training and inference, weight the FP8 and FP16 rows and memory capacity most heavily. For scientific and HPC workloads, the FP64 row is the one to read.

Frequently asked questions

Is the B100 faster than the GeForce RTX 5080?

At FP8 precision the B100 reaches 3,500 TFLOPS dense against 450 TFLOPS for the GeForce RTX 5080. The performance table on this page lists every published precision for both GPUs.

Which has more memory, the B100 or the GeForce RTX 5080?

The B100 carries 192 GB of VRAM versus 16 GB for the GeForce RTX 5080. Memory bandwidth is 8.0 TB/s for the B100 and 960 GB/s for the GeForce RTX 5080.

How much power do the B100 and the GeForce RTX 5080 draw?

The B100 is rated at 700 W TDP and the GeForce RTX 5080 at 360 W. On FP32 throughput per watt, the GeForce RTX 5080 is the more efficient part.

Can I rent the B100 or the GeForce RTX 5080 in the cloud?

Yes. Live cloud listings tracked by Flopper start at $0.39 per GPU hour. The rental pricing section on this page lists current providers and rates for both GPUs.

Where to Rent

NVIDIA GeForce RTX 5080 16GB

ProviderConfigurationPrice/GPU-hrChecked
RunPod
1× GeForce RTX 5080 16GB Community $0.39 5h ago View →
RunPod
1× GeForce RTX 5080 16GB $0.59 5h ago View →
All GeForce RTX 5080 listings and price history →

Get Comparison Updates

New GPUs added weekly. Be the first to see how they compare.