HW

Huawei Atlas 800T A3

rack
8× Ascend 910C
UnifiedBus (UB): 784 GB/s bidirectional D2D per NPU; scales to a 384-NPU Atlas 900 A3 SuperPoD
2025
FP16
6.0
PFLOPS
Power
kW Total
Memory
1049
GB Total

The training member of Huawei's Atlas 800 A3 family: a 10U air-cooled compute node with eight Ascend 910 NPUs and four Kunpeng 920 CPUs, rated at 6.0 PFLOPS of FP16 with 1,024 GB of on-chip memory at 3.2 TB/s per NPU and 784 GB/s of bidirectional device-to-device bandwidth. Huawei sells three models on this chassis at 6.0, 5.0 and 4.48 PFLOPS FP16, which works out at 750, 625 and 560 TFLOPS per NPU; this is the top one, and the 4.48 model is the separately listed Atlas 800I A3. Multiple nodes combine into an Atlas 900 A3 SuperPoD of up to 384 cards. Huawei publishes only an FP16 figure for this machine, and no INT8: a previously held INT8 value of 12.0 PFLOPS was 8 x 1,500 TFLOPS, a per-NPU rate Huawei states nowhere, and has been removed. Power is also not attributable: Huawei's family page gives 16.2 kW as the maximum input power for the Atlas 800 A3 line, but that line covers three compute variants and the Atlas 800I A3's own page separately says 14.6 kW, so no figure is recorded here rather than guessing which model the 16.2 kW describes. Huawei labels nothing dense or sparse.

We will point you at suppliers who have it. Free, and no signup.

FP16
6.00
PFLOPS

System Details

GPU Configuration

GPU Count: 8 GPUs
Architecture: Da Vinci v2 (dual-die)
Interconnect: UnifiedBus (UB): 784 GB/s bidirectional D2D per NPU; scales to a 384-NPU Atlas 900 A3 SuperPoD

System Specifications

Form Factor: rack
Total Power:
Total Memory: 1049 GB
Memory Bandwidth: 25600 GB/s

Precision Performance Breakdown

PrecisionSystem PerformancePer GPUEfficiency
6.000 PFLOPS 750.0 TFLOPS

All figures are dense. Vendors commonly headline the number, which is twice the dense one.

Powered by Huawei Ascend 910C

This system utilizes 8 × Huawei Ascend 910C GPUs, each delivering exceptional performance for AI and HPC workloads.

Per GPU TDP

600W

Per GPU Memory

128 GB

Process Node

SMIC 7nm

Architecture

Da Vinci v2 (dual-die)

Documentation & Resources

Official Datasheet

Huawei Atlas 800T A3 technical specifications

Download PDF

Atlas 650E 服务器 技术规格 (Huawei Ascend AI server product page)

Huawei • Latest version

View Document ↗

Typical Use Cases

AI/ML Training
High-Performance Computing
Data Analytics

The Huawei Atlas 800T A3 runs 8× Ascend 910C GPUs over UnifiedBus (UB): 784 GB/s bidirectional D2D per NPU; scales to a 384-NPU Atlas 900 A3 SuperPoD, delivering 6.00 PFLOPS FP16 dense.