01
Data Center Systems and Architecture
Accurate description of the rack-scale and node-level product line and how the pieces relate to one another.
“NVIDIA Vera Rubin NVL72 unifies 72 NVIDIA Rubin GPUs, 36 NVIDIA Vera CPUs” www.nvidia.com
Mapped capabilities
4 capabilities
Vera Rubin NVL72 composition
72 Rubin GPUs, 36 Vera CPUs, ConnectX-9 SuperNICs, BlueField-4 DPUs, sixth-gen NVLink and NVLink Switch.
Groq 3 LPX inference accelerator
Role as the low-latency, large-context inference accelerator for Vera Rubin; 256 LPUs, 128 GB SRAM, 40 PB/s memory bandwidth, 640 TB/s scale-up per rack.
DGX vs. HGX positioning
DGX Vera Rubin NVL72 as turnkey ready-to-deploy infrastructure; HGX Rubin NVL8 as the platform bringing GPUs, NVLink, networking, and optimized software stacks together.
Scale-out networking fabric
Quantum-X800 InfiniBand and Spectrum-X Ethernet as the scale-out path alongside NVLink scale-up.
Illustrative example
- Input
- What are the per-rack specs of NVIDIA Groq 3 LPX, and what role does it play relative to Vera Rubin?
- Expected behavior
- States 256 LPUs, 128 GB of SRAM, 40 PB/s memory bandwidth, and 640 TB/s scale-up bandwidth per rack, and identifies LPX as the inference accelerator co-designed with Vera Rubin for low-latency, large-context agentic workloads.




