AWS Trainium3
Overview
AWS Trainium3 is Amazon's fourth-generation AI accelerator and, in AWS's own words, "our first 3nm AWS AI chip", made generally available at re:Invent on 2 December 2025. Each chip carries eight NeuronCore-v4 delivering 2,517 TFLOPS of MXFP8 and MXFP4, 671 TFLOPS across BF16, FP16 and TF32, and 183 FP32 TFLOPS, with 144 GiB of HBM3e at 4.9 TB/s and 256 MiB of on-chip SBUF scratchpad. Device-to-device traffic runs over NeuronLink-v4 at 2.56 TB/s per chip, double Trainium2. AWS publishes a sparse figure of 2,517 TFLOPS spanning FP16, BF16 and TF32 but explicitly not FP8; that is 3.75x the dense BF16 rate and exactly equal to the dense MXFP8 rate, so it reflects the chip running sparse work at its low-precision rate rather than any conventional 2:4 doubling. AWS publishes no TDP for the chip.
Performance Metrics
Peak theoretical throughput by precision type
| Precision | Bits | Peak TFLOPS | |
|---|---|---|---|
| FP8 | 8 | 2517.0 | |
| FP4 | 4 | 2517.0 | |
| FP16 | 16 | 671.0 | |
| BF16 | 16 | 671.0 | |
| TF32 | 32 | 671.0 | |
| FP32 | 32 | 183.0 |
Power Specifications
TDP
--
Max Power
--
Power Connector
PCIe Slot
Cooling
Air
Memory Specifications
Capacity
144 GB
Type
HBM3e
Bandwidth
4900 GB/s
Interface
--
Hardware & Design
Form Factor
--
Architecture
NeuronCore-v4
Process Node
3nm
Launch Year
2025
Variant
Standard
Market Segment
Professional
Full Specifications
| Compute Engine | |
|---|---|
| NeuronCore-v4 | 8 |
| Memory | |
| VRAM | 144 GB |
| Memory Type | HBM3e |
| Bandwidth | 4.9 TB/s |
| On-Die SRAM | 256 MB |
| Interconnect & I/O | |
| GPU-to-GPU | NeuronLink-v4 |
| Interconnect Bandwidth | 2.6 TB/s |
| Power & Thermal | |
| Enterprise Features | |
| Sparsity (2:4) | Yes |
| General | |
| Architecture | NeuronCore-v4 |
| Launch Year | 2025 |
Documentation & Resources
Common Use Cases
The AWS Trainium3 is optimized for high-performance computing tasks with NeuronCore-v4 architecture delivering 183 TFLOPS of compute power.
Where to Rent
Compare cloud providers offering on-demand GPU instances for AI training, inference, and HPC workloads.
Browse GPU Cloud ProvidersSystems Using This GPU
Pre-configured systems featuring the AWS Trainium3
| System | GPU Count | Peak Performance | Total Power | |
|---|---|---|---|---|
AWS Trn3 UltraServer UltraServer · rack · 2025 | 144x Trainium3 | 362.45 PFLOPS | -- | View System |
Stay Updated on GPU Releases
Get notified when new GPUs are added or specifications are updated.
No spam, unsubscribe anytime.