NVIDIA GH200 96GB HBM3 vs NVIDIA GH200 144GB HBM3e

The GH200 96GB HBM3 and the GH200 144GB HBM3e deliver near-identical FP8 throughput (1,979 vs 1,979 TFLOPS dense). The GH200 144GB HBM3e also carries 48 GB more memory (144 GB vs 96 GB).

GH200 96GB HBM3: Hopper, 2023 GH200 144GB HBM3e: Hopper, 2024
VRAM
96 vs 144 GB
48 GB more for the GH200 144GB HBM3e
Bandwidth
4.0 TB/s vs 4.9 TB/s
GH200 144GB HBM3e moves data faster
TDP
1.0 kW vs 1.0 kW
Same power budget

Specifications

GH200 96GB HBM3
GH200 144GB HBM3e
Architecture
HopperHopper
Launch Year
20232024
Form Factor
SXMSXM
VRAM
96 GB144 GB
Memory Bandwidth
4.0 TB/s4.9 TB/s
TDP
1.0 kW1.0 kW
Process Node
4nm4nm

Performance (TFLOPS)

GH200 96GB HBM3
GH200 144GB HBM3e
FP64
34.0 TFlops 34.0 TFlops
FP32
67.0 TFlops 67.0 TFlops
FP16
990.0 TFlops
1979.0 TFLOPS sparse
990.0 TFlops
1979.0 TFLOPS sparse
BF16
990.0 TFlops
1979.0 TFLOPS sparse
990.0 TFlops
1979.0 TFLOPS sparse
FP8
1979.0 TFlops
3958.0 TFLOPS sparse
1979.0 TFlops
3958.0 TFLOPS sparse
INT8
1979.0 TFlops
3958.0 TOPS sparse
1979.0 TFlops
3958.0 TOPS sparse

FLOPS by Precision

What actually differs

The GH200 144GB HBM3e is the newer part: Hopper, launched in 2024, against the GH200 96GB HBM3's Hopper from 2023. Newer architectures typically add lower-precision formats and better throughput per watt, so check the precision rows your workload actually uses.

Dense throughput for the GH200 96GB HBM3 against the GH200 144GB HBM3e: FP64 34 vs 34 TFLOPS, FP32 67 vs 67 TFLOPS, FP16 990 vs 990 TFLOPS, FP8 1,979 vs 1,979 TFLOPS. Sparse figures, where the vendor publishes them, appear under each dense number in the performance table.

Memory is 96 GB against 144 GB, fed at 4.0 TB/s versus 4.9 TB/s. More memory per GPU means larger models fit before you have to shard across cards, which often matters more than raw TFLOPS for inference.

Power budgets are 1.0 kW for the GH200 96GB HBM3 and 1.0 kW for the GH200 144GB HBM3e. At FP32 that works out to 0.07 against 0.07 TFLOPS per watt.

For LLM training and inference, weight the FP8 and FP16 rows and memory capacity most heavily. For scientific and HPC workloads, the FP64 row is the one to read.

Frequently asked questions

Is the GH200 96GB HBM3 faster than the GH200 144GB HBM3e?

They are close on paper: both deliver about 1,979 TFLOPS of dense FP8 throughput. Memory, bandwidth, and power are the deciding differences.

Which has more memory, the GH200 96GB HBM3 or the GH200 144GB HBM3e?

The GH200 144GB HBM3e carries 144 GB of VRAM versus 96 GB for the GH200 96GB HBM3. Memory bandwidth is 4.0 TB/s for the GH200 96GB HBM3 and 4.9 TB/s for the GH200 144GB HBM3e.

How much power do the GH200 96GB HBM3 and the GH200 144GB HBM3e draw?

The GH200 96GB HBM3 is rated at 1.0 kW TDP and the GH200 144GB HBM3e at 1.0 kW.

Can I rent the GH200 96GB HBM3 or the GH200 144GB HBM3e in the cloud?

Yes. Live cloud listings tracked by Flopper start at $6.50 per GPU hour. The rental pricing section on this page lists current providers and rates for both GPUs.

Where to Rent

NVIDIA GH200 96GB HBM3

ProviderConfigurationPrice/GPU-hrChecked
CoreWeave
1× GH200 96GB HBM3 $6.50 2h ago View →
All GH200 96GB HBM3 listings and price history →

Get Comparison Updates

New GPUs added weekly. Be the first to see how they compare.