benchmarkdeploymentinfrastructurenvidia
NVIDIA Vera Rubin NVL72 Sets New Records in MLPerf Inference v6.1 Debut

Loading…

NVIDIA's Vera Rubin NVL72 system made its MLPerf Inference v6.1 debut and delivered leading performance across key AI inference benchmarks. The NVL72 configuration combines Vera CPUs with Rubin GPUs in a rack-scale design optimized for large model throughput. This is the first public benchmark validation of the Rubin architecture, giving developers and infrastructure teams concrete performance data to inform procurement and deployment decisions. For teams building or scaling inference infrastructure, the MLPerf results provide a direct comparison point against prior Hopper- and Blackwell-based systems. The results signal that NVIDIA's next-generation stack is production-ready and competitive at scale.