Meta MTIA 450 NEW
Overview
The Meta MTIA 450 is a generative AI inference specialist, and the point where Meta stopped optimising for balance and went after decode throughput. Meta doubled HBM bandwidth from the MTIA 400 to 18.4 TB/s, which it describes as much higher than existing leading commercial products, because memory bandwidth is the factor that most limits generative AI inference performance. HBM capacity stays at 288 GB and module power rises to 1400 W. Compute reaches 21 PFLOPS of MX4, 7 PFLOPS of FP8 and 3.5 PFLOPS of BF16, an increase of 75 percent in MX4 throughput over the MTIA 400. Beyond raw numbers Meta added hardware acceleration aimed at specific inference bottlenecks, easing Softmax and FlashAttention pressure and speeding up mixture-of-experts feed-forward network computation. It also supports mixed low-precision computation without the software overhead that data type conversion normally imposes, and introduces custom Meta data types intended to raise throughput while preserving model quality at minimal cost in chip area. Meta frames the balance plainly: the part delivers six times the MX4 throughput of its own FP16 and BF16, which is a statement about how far inference has moved toward low precision.
Performance Metrics
Peak theoretical throughput by precision type
| Precision | Bits | Peak TFLOPS | Efficiency |
|---|---|---|---|
| FP8 | 8 | 7000.0 | 5.000 TFLOPS/W |
| FP4 | 4 | 21000.0 | 15.000 TFLOPS/W |
| BF16 | 16 | 3500.0 | 2.500 TFLOPS/W |
Power Specifications
TDP
1.4 kW
Max Power
1.6 kW
Power Connector
PCIe 16-pin
Cooling
Air
Memory Specifications
Capacity
288 GB
Type
HBM
Bandwidth
18400 GB/s
Interface
--
Hardware & Design
Form Factor
--
Architecture
MTIA
Process Node
--
Launch Year
2027
Variant
Standard
Market Segment
Professional
Full Specifications
| Memory | |
|---|---|
| VRAM | 288 GB |
| Memory Type | HBM |
| Bandwidth | 18.4 TB/s |
| Interconnect & I/O | |
| Interconnect Bandwidth | 2.4 TB/s |
| Power & Thermal | |
| TDP | 1.4 kW |
| Enterprise Features | |
| Compute APIs | PyTorch, Triton |
| General | |
| Architecture | MTIA |
| Launch Year | 2027 |
Documentation & Resources
Common Use Cases
The Meta MTIA 450 is optimized for high-performance computing tasks with MTIA architecture delivering high TFLOPS of compute power.
Where to Rent
Compare cloud providers offering on-demand GPU instances for AI training, inference, and HPC workloads.
Browse GPU Cloud ProvidersSimilar GPUs
Stay Updated on GPU Releases
Get notified when new GPUs are added or specifications are updated.
No spam, unsubscribe anytime.