Qualcomm AI200
Overview
Qualcomm AI200 is a rack-scale AI inference accelerator announced on 27 October 2025, Qualcomm's return to the data center and the first product in its Dragonfly line, now branded Qualcomm Dragonfly AI200. Its defining decision is memory. Each card carries 768 GB of LPDDR5X, the highest per-accelerator memory capacity of any announced AI accelerator and more than five times what a 144 GB HBM3e part offers, deliberately trading bandwidth for capacity and cost so that very large models fit in fewer accelerators. Qualcomm has demonstrated a 350 billion parameter model running on a single card and states the card is designed for models up to a trillion parameters. Almost everything else Qualcomm has published is a rack aggregate rather than a card specification, and is recorded here as such: a rack holds 56 cards in a single-wide Open Rack v3 chassis for 43 TB of total memory and 0.414 PB/s of aggregate bandwidth, uses PCIe 6.0 for scale-up and Ethernet with RoCE for scale-out, supports confidential computing, and is offered with direct liquid cooling or air cooling. Qualcomm's launch materials quote rack-level power of 160 kW while its current product page lists 140 kW, and it has never published a per-card wattage, so no TDP is recorded here rather than dividing a rack figure across 56 cards. Most notably, Qualcomm has published no throughput figure of any kind, at any precision, at either card or rack level. There are no TOPS, no TFLOPS and no FP8, FP16 or INT8 numbers in any Qualcomm document, so this entry carries no performance metrics rather than estimated ones. That is a deliberate change in posture: for the predecessor Cloud AI 100 Ultra, Qualcomm published a product brief listing 150 W, 870 INT8 TOPS and 548 GB/s. Software runs through the Qualcomm AI Inference Suite and the Efficient Transformers Library with one-click deployment of Hugging Face models. The announced launch customer is HUMAIN of Saudi Arabia, targeting 200 MW of AI200 and AI250 rack deployments from 2026.
Performance
Peak theoretical throughput by precision type
| Precision | Peak |
|---|---|
FP64 | No verified data available |
FP32 | No verified data available |
TF32 | No verified data available |
BF16 | No verified data available |
FP16 | No verified data available |
FP8 | No verified data available |
FP6 | No verified data available |
FP4 | No verified data available |
INT8 | No verified data available |
Flopper holds no verified throughput figures for this part yet. The rows stay so the gaps are visible.
Specifications
Architecture
Dragonfly
Form Factor
No verified data available
Launch Year
2026
Process Node
No verified data available
Memory
768 GB LPDDR5X
Bandwidth
No verified data available
TDP
No verified data available
Max power
No verified data available
Interconnect
PCIe 6.0 scale-up / Ethernet with RoCE scale-out
Spec Confidence
Vendor claimed
Full Specifications
| Memory | |
|---|---|
| Memory | 768 GB |
| Memory Type | LPDDR5X |
| Bandwidth | No verified data available |
| Interface Width | No verified data available |
| Interconnect & I/O | |
| GPU-to-GPU | PCIe 6.0 scale-up / Ethernet with RoCE scale-out |
| Power & Thermal | |
| TDP | No verified data available |
| Max power | No verified data available |
| Cooling | Direct liquid cooling or air cooling |
| Enterprise Features | |
| Compute APIs | Qualcomm AI Inference Suite, Efficient Transformers Library, Hugging Face |
| General | |
| Form Factor | No verified data available |
| Architecture | Dragonfly |
| Process Node | No verified data available |
| Launch Year | 2026 |
Datasheet & Resources
Data Provenance
Every figure traced to a source
Primary Source
- Publisher
- Qualcomm
- Published
- No verified data available
Data Quality
- Spec confidence
- Vendor claimed
- Core precisions with figures
- 0 of 9
- Normalization
- All values in TFLOPS
Compare cloud providers offering on-demand GPU instances for AI training, inference, and HPC workloads.
Browse GPU Cloud ProvidersSimilar GPUs
Frequently Asked Questions
How many TFLOPS does the Qualcomm AI200 have?
Flopper does not currently have any verified throughput figures for the Qualcomm AI200.
What is the power consumption of the Qualcomm AI200?
Flopper does not currently have a verified TDP (Thermal Design Power) figure for the Qualcomm AI200.
How much memory does the Qualcomm AI200 have?
The Qualcomm AI200 is equipped with 768 GB of memory. Flopper does not currently have a verified memory bandwidth figure for it.
What architecture is the Qualcomm AI200 based on?
The Qualcomm AI200 is based on the Dragonfly architecture, launched in 2026.
Stay Updated on GPU Releases
Get notified when new GPUs are added or specifications are updated.
No spam, unsubscribe anytime.