Specifications

Rack Configuration

72x NVIDIA B200 GPU single open rack v3 solution

GPU Allocation

18x compute trays + 9x non-scalable NVLink switch trays

CPUs

36x NVIDIA Grace (2 per compute tray)

GPU Interconnect

NVIDIA NVLink 5, 130 TB/s aggregate per rack

GPU Memory

13.4 TB HBM3E pooled per rack

East-West Networking

NVIDIA ConnectX-7 / ConnectX-8

North-South Networking

FHFL NVIDIA BlueField-3 DPU

Rack Platform

440U OCP Open Rack v3 (ORv3)

Bus Bar

1400 A

Power Shelves

4x 66 kW

Cooling Support

Liquid-to-air and liquid-to-liquid CDU options

Rack Manager

Wiwynn UMS100 liquid cooling rack manager

Management API

Redfish, with GUI

Operating Power

Approximately 120 kW at rated NVFP4 workload

Target Workloads

Trillion-parameter training, MoE inference, reasoning models

Overview

The Wiwynn GB200 NVL72 is a rack-scale system rather than a server: the NVLink domain spans all 72 Blackwell GPUs, so the rack behaves as one accelerator with a 13.4 TB pool of HBM3E and 130 TB/s of all-to-all bandwidth. Wiwynn partitions it as eighteen 1U compute trays and nine NVLink switch trays inside a 440U OCP Open Rack v3.

Power delivery is a 1400 A bus bar fed by four 66 kW power shelves, which is why the platform is liquid cooled by design rather than as an option. Wiwynn supports both liquid-to-air and liquid-to-liquid CDU topologies, so the rack can be deployed in facilities without a chilled-water loop.

Networking is split cleanly: ConnectX-7 and ConnectX-8 handle east-west GPU-to-GPU traffic inside the rack, while FHFL BlueField-3 DPUs carry north-south ingress and egress. Cooling is monitored by the companion UMS100 rack manager, which exposes leak detection, telemetry and cooling optimisation over Redfish.

Key Benefits

Single NVLink domain: 130 TB/s of GPU-to-GPU bandwidth removes the cross-node penalty that limits large MoE and long-context inference. Deployment flexibility: liquid-to-air or liquid-to-liquid CDU support widens the set of facilities that can host the rack. Integrated management: Redfish-exposed bus bar, power shelf and cooling telemetry from the UMS100 rather than a bolt-on BMS.

Applications

Frontier model pre-training, mixture-of-experts inference at scale, long-context reasoning and agentic workloads, sovereign AI clusters, and hyperscale AI factory deployments.

Request a Quote — Wiwynn GB200 NVL72 Rack-Scale AI System — 72x NVIDIA B200 in One ORv3 Rack

QS Compute — global B2B supply of AI computing hardware, edge AI systems and accelerators. Volume pricing, 15-day sample lead time.

Get Your Quote →

Related Products

NVIDIA GB200 NVL72 NVIDIA B300 NVL8 SuperPOD Wiwynn UMS100 Cooling Manager NVIDIA ConnectX-9 SuperNIC