AW

AWS Trainium3

NeuronCore-v4 2025 3nm
FP32
183.0
TFLOPS
VRAM
144 GB
HBM3e
Bandwidth
4.9 TB/s
memory
NeuronCore-v4
8

Overview

AWS Trainium3 is Amazon's fourth-generation AI accelerator and, in AWS's own words, "our first 3nm AWS AI chip", made generally available at re:Invent on 2 December 2025. Each chip carries eight NeuronCore-v4 delivering 2,517 TFLOPS of MXFP8 and MXFP4, 671 TFLOPS across BF16, FP16 and TF32, and 183 FP32 TFLOPS, with 144 GiB of HBM3e at 4.9 TB/s and 256 MiB of on-chip SBUF scratchpad. Device-to-device traffic runs over NeuronLink-v4 at 2.56 TB/s per chip, double Trainium2. AWS publishes a sparse figure of 2,517 TFLOPS spanning FP16, BF16 and TF32 but explicitly not FP8; that is 3.75x the dense BF16 rate and exactly equal to the dense MXFP8 rate, so it reflects the chip running sparse work at its low-precision rate rather than any conventional 2:4 doubling. AWS publishes no TDP for the chip.

Performance Metrics

Peak theoretical throughput by precision type

PrecisionBitsPeak TFLOPS
FP8 8 2517.0
FP4 4 2517.0
FP16 16 671.0
BF16 16 671.0
TF32 32 671.0
FP32 32 183.0

Power Specifications

TDP

--

Max Power

--

Power Connector

PCIe Slot

Cooling

Air

Memory Specifications

Capacity

144 GB

Type

HBM3e

Bandwidth

4900 GB/s

Interface

--

Hardware & Design

Form Factor

--

Architecture

NeuronCore-v4

Process Node

3nm

Launch Year

2025

Variant

Standard

Market Segment

Professional

Full Specifications

Compute Engine
NeuronCore-v4 8
Memory
VRAM 144 GB
Memory Type HBM3e
Bandwidth 4.9 TB/s
On-Die SRAM 256 MB
Interconnect & I/O
GPU-to-GPU NeuronLink-v4
Interconnect Bandwidth 2.6 TB/s
Power & Thermal
Enterprise Features
Sparsity (2:4) Yes
General
Architecture NeuronCore-v4
Launch Year 2025

Documentation & Resources

Common Use Cases

General Compute AI/ML Workloads Data Processing

The AWS Trainium3 is optimized for high-performance computing tasks with NeuronCore-v4 architecture delivering 183 TFLOPS of compute power.

Where to Rent

Compare cloud providers offering on-demand GPU instances for AI training, inference, and HPC workloads.

Browse GPU Cloud Providers

Systems Using This GPU

Pre-configured systems featuring the AWS Trainium3

SystemGPU CountPeak PerformanceTotal Power
AWS Trn3 UltraServer
UltraServer · rack · 2025
144x Trainium3 362.45 PFLOPS -- View System

Stay Updated on GPU Releases

Get notified when new GPUs are added or specifications are updated.

Loading verification...

No spam, unsubscribe anytime.

Back to GPUs