MPN: RUBIN-R100
NVIDIA Rubin GPU
288GB HBM4 · 50 PFLOPS NVFP4 · 336B Transistors

Rubin Architecture · 3rd-gen Transformer Engine · NVLink 6 · TSMC 3nm · Full Production

Specifications

Model

NVIDIA Rubin GPU (R100)

Architecture

NVIDIA Rubin
3rd-gen Transformer Engine

Transistors

336 billion

VRAM

288 GB HBM4

Memory Bandwidth

22 TB/s per GPU

FP4 Inference

50 PFLOPS (NVFP4)

FP4 Training

35 PFLOPS

NVLink

NVLink 6
3.6 TB/s per GPU

Process

TSMC 3nm (3NP/3PN)

HBM4

10.8 GT/s per pin
>3.0 TB/s per stack

Paired CPU

NVIDIA Vera
88 Olympus cores · Armv9.2

Networking

ConnectX-9 SuperNIC
1.6 Tb/s per GPU

Confidential Computing

3rd-generation

RAS

2nd-gen RAS engine

Availability

Full production · partners H2 2026

Overview

The NVIDIA Rubin GPU is the successor to Blackwell and the compute heart of the Vera Rubin platform, named for the pioneering astronomer Vera Rubin. With 336 billion transistors, 288 GB of HBM4 at 22 TB/s of memory bandwidth, and 50 petaflops of NVFP4 inference compute, Rubin sets a new bar for AI training and inference at gigascale.

Its third-generation Transformer Engine adds hardware-accelerated adaptive compression for efficient low-precision inference, while HBM4 delivers a 2.75× bandwidth improvement over Blackwell's HBM3e at equivalent capacity. Rubin GPUs pair with the NVIDIA Vera CPU — 88 custom Olympus cores with full Armv9.2 compatibility connected over ultra-fast NVLink-C2C — to form the Vera Rubin Superchip.

At the system level, Rubin powers the Vera Rubin NVL72 rack (72 GPUs + 36 Vera CPUs, 3.6 exaFLOPS FP4), the HGX Rubin NVL8 server board for x86 platforms, and DGX SuperPOD reference architectures. NVIDIA confirmed Rubin entered full production in early 2026, with partner systems shipping in the second half of the year.

Request a Quote — NVIDIA Rubin GPU

QS Compute — global B2B supply for NVIDIA data center and edge AI hardware. Enterprise procurement, direct manufacturer channels.

Request Quote