MPN: Dragonfly AI300
Qualcomm Dragonfly AI300
HBC Gen 2 · Rack-Scale Inference

Third-generation rack-level AI inference platform · integrated Qualcomm High Bandwidth Compute (HBC) Gen 2 · air and direct-liquid cooled · 4–8× better performance per watt versus GPU-based architectures

Dragonfly AI300

Third-generation, air- and direct-liquid-cooled rack-level AI inference platform, following AI200 and AI250 (October 2025).

AI300 Card

Accelerator card form factor for scale-out inference deployments inside standard racks.

AI300 Rack

Rack-scale inference system integrating accelerator, memory and networking into one platform.

AI200 / AI250 / AI300 Roadmap

Qualcomm's annual-cadence inference accelerator line, now in its third generation.

Specifications

Vendor

Qualcomm Technologies, Inc.

Product

Qualcomm Dragonfly AI300

Type

Rack-scale AI inference accelerator

Generation

Third generation (AI200 → AI250 → AI300)

Memory Architecture

Qualcomm High Bandwidth Compute (HBC) Gen 2

Efficiency

4–8× better performance per watt vs GPU-based architectures

Workloads

LLM, LMM, agentic AI, long-context inference

Deployment

Disaggregated inference

Scale-Up

UALink (Ultra Accelerator Link), ESUN (Ethernet for Scale-Up Networking)

Scale-Out

Copper and optical networking

Cooling

Air and direct-liquid cooling

Form Factors

Card and rack

Cadence

Annual accelerator roadmap

Availability

Commercial sampling expected 2028

Stock

Available

Price

Quote

Overview

The Qualcomm Dragonfly AI300 is the third generation of Qualcomm's data center inference accelerator line, announced on June 24, 2026 at the company's Investor Day alongside the Dragonfly C1000 CPU and Qualcomm High Bandwidth Compute (HBC). It succeeds the AI200 and AI250 launched in October 2025 on an annual cadence.

The AI300 integrates HBC Gen 2 — Qualcomm's 3D-stacked near-memory computing architecture — to pair compute with dramatically increased effective memory bandwidth. Qualcomm positions it for disaggregated inference deployments, targeting 4–8× better performance per watt than existing GPU-based architectures on a memory-bandwidth-per-watt-per-card basis, with industry-leading memory capacity for LLM, LMM and agentic AI serving.

Connectivity is built for modern AI factories: UALink and ESUN for scale-up, plus copper and optical scale-out. The AI300 ships in both card and rack form factors with air and direct-liquid cooling. Commercial sampling is expected in 2028.

Need Qualcomm Dragonfly AI300?

Contact QS Compute for availability, configuration, and volume pricing.

Request Quote