Specifications

GPUs

8× NVIDIA Blackwell Ultra SXM

GPU Memory

2.1TB HBM3e total

FP4 Performance

144 PFLOPS (sparse), 108 PFLOPS (dense)

FP8 Performance

72 PFLOPS (sparse)

NVLink

2× NVSwitch, 14.4 TB/s aggregate

Networking

8× ConnectX-8 800Gb/s InfiniBand/Ethernet

CPU

Intel Xeon 6776P Processors

Rack Compatibility

NVIDIA MGX + traditional enterprise racks

AI Reasoning

1.5× FP4 dense + 2× attention vs DGX B200

Cooling

Liquid-cooled

Use Case

LLM training, AI reasoning, inference serving

Overview

The NVIDIA DGX B300 is the first DGX system built on NVIDIA Blackwell Ultra architecture, delivering 1.5× the dense FP4 performance and 2× the attention performance of its predecessor DGX B200. With 8× Blackwell Ultra GPUs, 2.1TB of HBM3e memory, and 14.4 TB/s of GPU interconnect bandwidth, the DGX B300 is purpose-built for the AI reasoning era — enabling enterprises to run trillion-parameter models, agentic AI workflows, and high-throughput inference serving at hyperscale efficiency. First DGX system deployable in NVIDIA MGX racks, setting a new standard for data center AI infrastructure.

Key Benefits

AI Reasoning Engine: Purpose-built for LLM inference and training workloads that demand massive attention compute — 2× attention throughput vs DGX B200 enables real-time reasoning for agentic AI. MGX-Compatible: First DGX deployable in NVIDIA MGX modular racks alongside traditional enterprise racks, simplifying procurement and deployment. Sustainable Performance: Multiple power configuration options deliver highest performance-per-watt in its class. Full-Stack Software: NVIDIA AI Enterprise, NIM microservices, and CUDA-X libraries included — deploy AI factories in days, not months.

Request a Quote — NVIDIA DGX B300

QS Compute — NVIDIA Goldendisk Partner. Global B2B supply, enterprise procurement, direct manufacturer channels.

Get Your Quote →