NVIDIA Groq 3 LPU

Type: LPUArchitecture: Groq LPUReleased: 2026Spec confidence: Official
FP8 (dense)
1,200
TFLOPS

Overview

The Groq 3 LPU is the accelerator inside NVIDIA's Groq 3 LPX inference rack, and is not sold separately: it is soldered eight to a 1U tray, thirty-two trays to a rack, and there is no standalone SKU. It has no HBM, no GDDR and no hardware-managed cache hierarchy at all. Its working storage is 500 MB of on-die SRAM running at 150 TB/s, with 2.5 TB/s of chip-to-chip scale-up bandwidth, and it relies on a statically compiler-scheduled deterministic dataflow rather than caches. NVIDIA publishes 9.6 PFLOPS of FP8 per eight-LPU tray, which is 1,200 TFLOPS each, and that reconciles against the rack figure of 307.2 PFLOPS across 256 LPUs. The per-chip memory figures reconcile the same way: 256 x 500 MB is the rack's published 128 GB of SRAM, and 256 x 2.5 TB/s is its published 640 TB/s of scale-up bandwidth. Figures are dense: NVIDIA describes the matrix execution modules as providing dense multiply-accumulate capability, and a statically scheduled design has no structured-sparsity hardware, so no sparse rate exists.

Performance

Peak theoretical throughput by precision type

PrecisionPeak
FP64
No verified data available
FP32
No verified data available
TF32
No verified data available
BF16
No verified data available
FP16
No verified data available
FP8
8-bit floating point
1,200TFLOPS
FP6
No verified data available
FP4
No verified data available
INT8
No verified data available

Specifications

Architecture

Groq LPU

Form Factor

No verified data available

Launch Year

2026

Process Node

No verified data available

Memory

No verified data available On-die SRAM

Bandwidth

No verified data available

TDP

No verified data available

Max power

No verified data available

Interconnect

2.5 TB/s LPU C2C

direction and scope not stated by vendor

Spec Confidence

Official

Full Specifications

Memory
Memory No verified data available
Memory Type On-die SRAM
Bandwidth No verified data available
Interface Width No verified data available
On-Die SRAM 500 MB
Interconnect & I/O
GPU-to-GPU LPU C2C
Interconnect Bandwidth 2.5 TB/s direction and scope not stated by vendor
Power & Thermal
TDP No verified data available
Max power No verified data available
General
Form Factor No verified data available
Architecture Groq LPU
Process Node No verified data available
Launch Year 2026

Datasheet & Resources

Flopper Datasheet

NVIDIA Groq 3 LPU specifications, generated from our database. Printable.

Vendor product page

NVIDIA documentation for the Groq 3 LPU

View

NVIDIA Developer Documentation

Technical specs, programming guides

Browse

Data Provenance

Every figure traced to a source

Primary Source

Document
—
Publisher
NVIDIA
Published
No verified data available

Data Quality

Spec confidence
Official
Clock basis
Boost
Core precisions with figures
1 of 9
Normalization
All values in TFLOPS

Compare cloud providers offering on-demand GPU instances for AI training, inference, and HPC workloads.

Browse GPU Cloud Providers

Systems Using This GPU

Pre-configured systems featuring the NVIDIA Groq 3 LPU

SystemGPU CountPeak PerformanceTotal Power
NVIDIA Groq 3 LPX
LPX · rack · 2026
256x Groq 3 LPU 307 PFLOPS No verified data available View System

Frequently Asked Questions

How many TFLOPS does the NVIDIA Groq 3 LPU have?

The NVIDIA Groq 3 LPU delivers 1,200 TFLOPS FP8 at peak. Flopper does not currently have verified FP32 and FP16 throughput figures for it.

What is the power consumption of the NVIDIA Groq 3 LPU?

Flopper does not currently have a verified TDP (Thermal Design Power) figure for the NVIDIA Groq 3 LPU.

How much memory does the NVIDIA Groq 3 LPU have?

Flopper does not currently have verified memory capacity or bandwidth figures for the NVIDIA Groq 3 LPU.

What architecture is the NVIDIA Groq 3 LPU based on?

The NVIDIA Groq 3 LPU is based on the Groq LPU architecture, launched in 2026.

Stay Updated on GPU Releases

Get notified when new GPUs are added or specifications are updated.

Loading verification...

No spam, unsubscribe anytime.

Back to GPUs

© 2026 Flopper.io - Compare the GPUs Powering AI