Skip to content
NVIDIA Blog· Zhihan Jiang·· 15 d ago

NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut

NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut

AI summary

NVIDIA submitted preview results for Vera Rubin NVL72 to MLPerf Inference v6.1 for the first time, achieving throughput up to 3.7 times that of GB300 NVL72 on Qwen3-VL and up to 2.5 times on DeepSeek-R1.

Selection record

AdmittedSum of both 124 ≥ twice the threshold 120

Source tier
Official, first-hand; this tier's threshold is 60
Pre-filter
passed:NVIDIA AI推理平台MLPerf性能结果
Why it was chosen
The throughput and scaling-efficiency figures from its first MLPerf inference submission inform assessment of value-for-money trends in inference infrastructure.

A model scores each item twice, independently, against one written standard, out of 100. An item is admitted only when the two scores add up to twice the threshold. The threshold is set per source tier.

Source: NVIDIA Blog · blogs.nvidia.com