NVIDIA GH200 144GB HBM3e vs NVIDIA GH200 96GB HBM3

The GH200 144GB HBM3e and the GH200 96GB HBM3 deliver near-identical FP8 throughput (1,979 vs 1,979 TFLOPS dense). The GH200 144GB HBM3e also carries 48 GB more memory (144 GB vs 96 GB).

GH200 144GB HBM3e: Hopper, 2024 GH200 96GB HBM3: Hopper, 2023
Memory
144 vs 96 GB
48 GB more for the GH200 144GB HBM3e
Bandwidth
4.9 TB/s vs 4.0 TB/s
GH200 144GB HBM3e moves data faster
TDP
1.0 kW vs 1.0 kW
Same power budget

Specifications

GH200 144GB HBM3e
GH200 96GB HBM3
Architecture
HopperHopper
Launch Year
20242023
Form Factor
SuperchipSuperchip
Memory
144 GB96 GB
Memory Bandwidth
4.9 TB/s4.0 TB/s
TDP
1.0 kW1.0 kW
Process Node
4nm4nm

Performance (TFLOPS)

GH200 144GB HBM3e
GH200 96GB HBM3
FP64
34 TFLOPS 34 TFLOPS
FP32
67 TFLOPS 67 TFLOPS
TF32
494 TFLOPS
989 TFLOPS sparse
494 TFLOPS
989 TFLOPS sparse
BF16
990 TFLOPS
1,979 TFLOPS sparse
990 TFLOPS
1,979 TFLOPS sparse
FP16
990 TFLOPS
1,979 TFLOPS sparse
990 TFLOPS
1,979 TFLOPS sparse
FP8
1,979 TFLOPS
3,958 TFLOPS sparse
1,979 TFLOPS
3,958 TFLOPS sparse
FP6
No verified data available No verified data available
FP4
No verified data available No verified data available
INT8
1,979 TOPS
3,958 TOPS sparse
1,979 TOPS
3,958 TOPS sparse
Tensor Core
FP64 (TC)
67 TFLOPS 67 TFLOPS

FLOPS by Precision

What actually differs

The GH200 144GB HBM3e is the newer part: Hopper, launched in 2024, against the GH200 96GB HBM3's Hopper from 2023. Newer architectures typically add lower-precision formats and better throughput per watt, so check the precision rows your workload actually uses.

Dense throughput for the GH200 144GB HBM3e against the GH200 96GB HBM3: FP64 34 vs 34 TFLOPS, FP32 67 vs 67 TFLOPS, FP16 990 vs 990 TFLOPS, FP8 1,979 vs 1,979 TFLOPS. Sparse figures, where the vendor publishes them, appear under each dense number in the performance table.

Memory is 144 GB against 96 GB, fed at 4.9 TB/s versus 4.0 TB/s. More memory per GPU means larger models fit before you have to shard across cards, which often matters more than raw TFLOPS for inference.

Power budgets are 1.0 kW for the GH200 144GB HBM3e and 1.0 kW for the GH200 96GB HBM3. At FP32 that works out to 0.07 against 0.07 TFLOPS per watt.

For LLM training and inference, weight the FP8 and FP16 rows and memory capacity most heavily. For scientific and HPC workloads, the FP64 row is the one to read.

Frequently asked questions

Is the GH200 144GB HBM3e faster than the GH200 96GB HBM3?

They are close on paper: both deliver about 1,979 TFLOPS of dense FP8 throughput. Memory, bandwidth, and power are the deciding differences.

Which has more memory, the GH200 144GB HBM3e or the GH200 96GB HBM3?

The GH200 144GB HBM3e carries 144 GB of memory versus 96 GB for the GH200 96GB HBM3. Memory bandwidth is 4.9 TB/s for the GH200 144GB HBM3e and 4.0 TB/s for the GH200 96GB HBM3.

How much power do the GH200 144GB HBM3e and the GH200 96GB HBM3 draw?

The GH200 144GB HBM3e is rated at 1.0 kW TDP and the GH200 96GB HBM3 at 1.0 kW.

Can I rent the GH200 144GB HBM3e or the GH200 96GB HBM3 in the cloud?

Yes. Live cloud listings tracked by Flopper start at $2.29 per GPU hour. The rental pricing section on this page lists current providers and rates for both GPUs.

Where to Rent

NVIDIA GH200 96GB HBM3

ProviderConfigurationPrice/GPU-hrChecked
1× GH200 96GB HBM3 $2.29 2h ago View →
Spheron
1× GH200 96GB HBM3 $2.75 2h ago View →
1× GH200 96GB HBM3 $6.50 2h ago View →
Shadeform Latitude.sh capacity
1× GH200 96GB HBM3 $4.23 2h ago View →
All GH200 96GB HBM3 listings and price history →

Get Comparison Updates

New GPUs added weekly. Be the first to see how they compare.

© 2026 Flopper.io - Compare the hardware powering AI