NVIDIA Quadro P4000 8GB vs NVIDIA Tesla P40 24GB
The Tesla P40 delivers 2.3x the FP32 throughput of the Quadro P4000 (12 vs 5.3 TFLOPS dense). The Tesla P40 also carries 16 GB more memory (24 GB vs 8 GB).
NVIDIA Quadro P4000 8GB
Full specs →NVIDIA Tesla P40 24GB
Full specs →Specifications
Performance (TFLOPS)
FLOPS by Precision
What actually differs
The Quadro P4000 is the newer part: Pascal, launched in 2017, against the Tesla P40's Pascal from 2016. Newer architectures typically add lower-precision formats and better throughput per watt, so check the precision rows your workload actually uses.
Dense throughput for the Quadro P4000 against the Tesla P40: FP32 5.3 vs 12 TFLOPS. Sparse figures, where the vendor publishes them, appear under each dense number in the performance table.
Memory is 8 GB against 24 GB, fed at 243 GB/s versus 346 GB/s. More memory per GPU means larger models fit before you have to shard across cards, which often matters more than raw TFLOPS for inference.
Power budgets are 105 W for the Quadro P4000 and 250 W for the Tesla P40. At FP32 that works out to 0.05 against 0.05 TFLOPS per watt.
For LLM training and inference, weight the FP8 and FP16 rows and memory capacity most heavily. For scientific and HPC workloads, the FP64 row is the one to read.
Frequently asked questions
Is the Quadro P4000 faster than the Tesla P40?
At FP32 precision the Tesla P40 reaches 12 TFLOPS dense against 5.3 TFLOPS for the Quadro P4000. The performance table on this page lists every published precision for both GPUs.
Which has more memory, the Quadro P4000 or the Tesla P40?
The Tesla P40 carries 24 GB of memory versus 8 GB for the Quadro P4000. Memory bandwidth is 243 GB/s for the Quadro P4000 and 346 GB/s for the Tesla P40.
How much power do the Quadro P4000 and the Tesla P40 draw?
The Quadro P4000 is rated at 105 W TDP and the Tesla P40 at 250 W. On FP32 throughput per watt, the Quadro P4000 is the more efficient part.
Get Comparison Updates
New GPUs added weekly. Be the first to see how they compare.