Specifications
NVIDIA Part Number
900-5G153-2200-000
PNY ordering part
VCNRTXPRO6000BQ-PB
Product
NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition
GPU
GB202 (Blackwell)
GPU Memory
96 GB GDDR7 with ECC
Memory Interface
512-bit
Memory Bandwidth
Up to 1,792 GB/s
NVIDIA CUDA Cores
24,064
Tensor Cores
752 (5th generation)
RT Cores
188 (4th generation)
FP32 Single Precision
110 TFLOPS (peak)
FP4 AI (with sparsity)
3.511 PFLOPS (peak)
RT Core Performance
333 TFLOPS (peak)
Host Interface
PCIe 5.0 x16
NVLink
Not supported
Power Consumption
300 W
Cooling
Single-fan blower, standard height, dual slot
Form Factor Dimensions
4.4 in (H) × 10.5 in (L)
Display Outputs
4 × DisplayPort 2.1
Lenovo ordering part
4X67B09095 (feature CBU5)
Overview
The NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition (part 900-5G153-2200-000) is the 300 W member of the RTX PRO 6000 Blackwell family. It carries the same GB202 GPU and the same 96 GB of GDDR7 ECC as the 600 W Workstation Edition, but is power-capped to 300 W with a single-fan blower in a standard-height dual-slot body.
That power envelope is the point of the Max-Q variant. A 600 W card dictates chassis airflow, PSU sizing and slot spacing; a 300 W card fits an ordinary workstation or a 2U-adjacent enclosure, and lets a system carry more cards per kilowatt. The cost is clock headroom - peak FP32 drops from 125 TFLOPS (Workstation Edition) to 110 TFLOPS, and FP4 sparsity from a higher figure to 3.511 PFLOPS.
What is preserved is capacity and connectivity: 96 GB of GDDR7 at 1,792 GB/s over a 512-bit bus, 24,064 CUDA cores, 752 fifth-generation Tensor cores and 188 fourth-generation RT cores. For workloads that are memory-capacity bound - large-model inference, Omniverse scenes, scientific simulation - that combination buys more usable throughput per rack unit than an uncapped card that cannot be densely deployed. Note the Max-Q part carries a controlled-GPU export status per Lenovo's ordering documentation.
Key Benefits
96 GB GDDR7 ECC at 300 W — full-capacity Blackwell in a deployable power budget. 3.511 PFLOPS FP4 — fifth-generation Tensor cores for AI inference. 1,792 GB/s memory bandwidth — 512-bit GDDR7 interface. Standard-height dual-slot blower — fits dense multi-GPU workstations.
Applications
Large-model inference and fine-tuning at the workstation tier, NVIDIA Omniverse and digital-twin simulation, professional rendering and ray tracing, scientific and engineering simulation, and multi-GPU deskside AI systems.
Request a Quote — 900-5G153-2200-000
QS Compute — global B2B supply of AI computing hardware, edge AI systems and accelerators. Volume pricing, 15-day sample lead time.
Get Your Quote →