Live Feed/NVIDIA/Fact Record
NVIDIA logo
NVIDIA
product launch 96% Confidence Gate September 16, 2026

NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut

NVIDIA has introduced the Vera Rubin NVL72 system, which achieved benchmark results in the MLPerf Inference v6.1 suite. This system represents the debut of the Vera Rubin architecture in standardized industry performance testing.

Verified State Diff

Comparison Mode:
- Previous State
Inference performance benchmarks were limited to Blackwell and Hopper-based architectures in previous MLPerf cycles.
+ Verified New State
The Vera Rubin NVL72 architecture is now officially benchmarked and available for performance evaluation in MLPerf Inference v6.1.

Impact & Verification Analysis

WHO IS AFFECTED

Enterprise AI infrastructure architects, cloud service providers, and large-scale model developers.

WHY IT MATTERS

It establishes the performance baseline for the next generation of NVIDIA hardware, directly impacting the cost-per-token economics for massive AI inference workloads.

Full Fact Overview

The Vera Rubin NVL72 leverages the next-generation Vera Rubin architecture, focusing on high-density inference throughput and scalable infrastructure. By participating in MLPerf Inference v6.1, NVIDIA provides verifiable data points on token generation rates and scaling efficiency for large-scale AI deployments. This release signals a shift toward the Rubin-based hardware cycle, emphasizing the economic necessity of maximizing tokens-per-watt and infrastructure utilization in data center environments.

Multi-Source Evidence Chain (1)

NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 DebutNVIDIA
TRACKED ENTITY
Explore all historical NVIDIA changes
View NVIDIA Hub ➔