NVIDIA GeForce GTX 1070 8GB vs NVIDIA P100 SXM2 16GB
The P100 delivers 64% more FP32 throughput than the GeForce GTX 1070 (11 vs 6.5 TFLOPS dense). The P100 also carries 8 GB more memory (16 GB vs 8 GB).
NVIDIA GeForce GTX 1070 8GB
Full specs →NVIDIA P100 SXM2 16GB
Full specs →Specifications
Performance (TFLOPS)
FLOPS by Precision
What actually differs
Both GPUs launched in 2016: the GeForce GTX 1070 on Pascal and the P100 on Pascal.
Dense throughput for the GeForce GTX 1070 against the P100: FP32 6.5 vs 11 TFLOPS. Sparse figures, where the vendor publishes them, appear under each dense number in the performance table.
Memory is 8 GB against 16 GB, fed at 256 GB/s versus 732 GB/s. More memory per GPU means larger models fit before you have to shard across cards, which often matters more than raw TFLOPS for inference.
Power budgets are 150 W for the GeForce GTX 1070 and 300 W for the P100. At FP32 that works out to 0.04 against 0.04 TFLOPS per watt.
For LLM training and inference, weight the FP8 and FP16 rows and memory capacity most heavily. For scientific and HPC workloads, the FP64 row is the one to read.
Frequently asked questions
Is the GeForce GTX 1070 faster than the P100?
At FP32 precision the P100 reaches 11 TFLOPS dense against 6.5 TFLOPS for the GeForce GTX 1070. The performance table on this page lists every published precision for both GPUs.
Which has more memory, the GeForce GTX 1070 or the P100?
The P100 carries 16 GB of memory versus 8 GB for the GeForce GTX 1070. Memory bandwidth is 256 GB/s for the GeForce GTX 1070 and 732 GB/s for the P100.
How much power do the GeForce GTX 1070 and the P100 draw?
The GeForce GTX 1070 is rated at 150 W TDP and the P100 at 300 W. On FP32 throughput per watt, the GeForce GTX 1070 is the more efficient part.
Get Comparison Updates
New GPUs added weekly. Be the first to see how they compare.