NVIDIA Groq 3 LPU
Overview
The Groq 3 LPU is the accelerator inside NVIDIA's Groq 3 LPX inference rack, and is not sold separately: it is soldered eight to a 1U tray, thirty-two trays to a rack, and there is no standalone SKU. It has no HBM, no GDDR and no hardware-managed cache hierarchy at all. Its working storage is 500 MB of on-die SRAM running at 150 TB/s, with 2.5 TB/s of chip-to-chip scale-up bandwidth, and it relies on a statically compiler-scheduled deterministic dataflow rather than caches. NVIDIA publishes 9.6 PFLOPS of FP8 per eight-LPU tray, which is 1,200 TFLOPS each, and that reconciles against the rack figure of 307.2 PFLOPS across 256 LPUs. The per-chip memory figures reconcile the same way: 256 x 500 MB is the rack's published 128 GB of SRAM, and 256 x 2.5 TB/s is its published 640 TB/s of scale-up bandwidth. Figures are dense: NVIDIA describes the matrix execution modules as providing dense multiply-accumulate capability, and a statically scheduled design has no structured-sparsity hardware, so no sparse rate exists.
Performance
Peak theoretical throughput by precision type
| Precision | Peak |
|---|---|
FP64 | No verified data available |
FP32 | No verified data available |
TF32 | No verified data available |
BF16 | No verified data available |
FP16 | No verified data available |
FP8 8-bit floating point | 1,200TFLOPS |
FP6 | No verified data available |
FP4 | No verified data available |
INT8 | No verified data available |
Specifications
Architecture
Groq LPU
Form Factor
No verified data available
Launch Year
2026
Process Node
No verified data available
Memory
No verified data available On-die SRAM
Bandwidth
No verified data available
TDP
No verified data available
Max power
No verified data available
Interconnect
2.5 TB/s LPU C2C
direction and scope not stated by vendor
Spec Confidence
Official
Full Specifications
| Memory | |
|---|---|
| Memory | No verified data available |
| Memory Type | On-die SRAM |
| Bandwidth | No verified data available |
| Interface Width | No verified data available |
| On-Die SRAM | 500 MB |
| Interconnect & I/O | |
| GPU-to-GPU | LPU C2C |
| Interconnect Bandwidth | 2.5 TB/s direction and scope not stated by vendor |
| Power & Thermal | |
| TDP | No verified data available |
| Max power | No verified data available |
| General | |
| Form Factor | No verified data available |
| Architecture | Groq LPU |
| Process Node | No verified data available |
| Launch Year | 2026 |
Datasheet & Resources
Flopper Datasheet
NVIDIA Groq 3 LPU specifications, generated from our database. Printable.
Vendor product page
NVIDIA documentation for the Groq 3 LPU
NVIDIA Developer Documentation
Technical specs, programming guides
Data Provenance
Every figure traced to a source
Primary Source
- Document
- —
- Publisher
- NVIDIA
- Published
- No verified data available
Data Quality
- Spec confidence
- Official
- Clock basis
- Boost
- Core precisions with figures
- 1 of 9
- Normalization
- All values in TFLOPS
Compare cloud providers offering on-demand GPU instances for AI training, inference, and HPC workloads.
Browse GPU Cloud ProvidersSystems Using This GPU
Pre-configured systems featuring the NVIDIA Groq 3 LPU
| System | GPU Count | Peak Performance | Total Power | |
|---|---|---|---|---|
NVIDIA Groq 3 LPX LPX · rack · 2026 | 256x Groq 3 LPU | 307 PFLOPS | No verified data available | View System |
Frequently Asked Questions
How many TFLOPS does the NVIDIA Groq 3 LPU have?
The NVIDIA Groq 3 LPU delivers 1,200 TFLOPS FP8 at peak. Flopper does not currently have verified FP32 and FP16 throughput figures for it.
What is the power consumption of the NVIDIA Groq 3 LPU?
Flopper does not currently have a verified TDP (Thermal Design Power) figure for the NVIDIA Groq 3 LPU.
How much memory does the NVIDIA Groq 3 LPU have?
Flopper does not currently have verified memory capacity or bandwidth figures for the NVIDIA Groq 3 LPU.
What architecture is the NVIDIA Groq 3 LPU based on?
The NVIDIA Groq 3 LPU is based on the Groq LPU architecture, launched in 2026.
Stay Updated on GPU Releases
Get notified when new GPUs are added or specifications are updated.
No spam, unsubscribe anytime.