Inspur NF5688M6
Inspur's six-rack-unit AI server, sold by its server arm IEIT SYSTEMS, and one of the few OEM machines whose maker names both the accelerator and the reference baseboard it sits on. It carries a single NVIDIA HGX A800 8-GPU baseboard holding eight SXM4 A800 modules, identified in Inspur's bill of materials as NV_640G_HGX-A800-8GPU, which places 640 GB of GPU memory in the chassis. NVSwitch gives any two GPUs a direct 400 GB/s path, and Inspur balances the machine at one InfiniBand port and one NVMe drive per GPU, with up to ten 100G or 200G RDMA adapters. Hosting is two third-generation Intel Xeon Scalable processors on the C621A chipset. Fully configured the machine weighs 88 kg net and 126 kg shipped, and it draws from six 3,000 W supplies in 3+3 redundancy. Inspur quotes 5 PFLOPS of AI performance for the machine but names no precision anywhere in the 89-page white paper and never uses the words dense or sparse, so no throughput figure is listed here. Inspur also declines to publish a system power figure, stating that consumption varies by configuration.
We will point you at suppliers who have it. Free, and no signup.
System Details
GPU Configuration
System Specifications
Powered by NVIDIA A800
This system utilizes 8 × NVIDIA A800 SXM4 GPUs, each delivering exceptional performance for AI and HPC workloads.
Per GPU TDP
—
Per GPU Memory
—
Process Node
—
Architecture
—
Documentation & Resources
Official Datasheet
Inspur NF5688M6 technical specifications
浪潮信息英信服务器 NF5688M6 技术白皮书 V1.6
Inspur • 2023-09-06
Typical Use Cases
The Inspur NF5688M6 runs 8× A800 GPUs over One NVIDIA HGX A800 8-GPU baseboard, with NVSwitch giving every pair of GPUs a direct 400 GB/s peer-to-peer path; up to ten 100G or 200G RDMA adapters for scale-out, in a 1:1:1 ratio of GPU to InfiniBand port to NVMe drive.