NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut
Article image or reusable cover for NVIDIA
NVIDIA presents its new Vera Rubin NVL72 system, which shows significant performance improvements in MLPerf Inference v6.1 tests.
The system delivers up to 3.7x higher throughput than its predecessor GB300 NVL72 on the Qwen3-VL benchmark, and 2.5x higher on DeepSeek-R1. When four GB300 NVL72 racks were scaled up together, 99 percent scaling efficiency was achieved — throughput grew nearly linearly as more GPUs were added. NVIDIA also achieved performance gains through software optimization: version 6.1 was up to 1.6x faster than 6.0. For AI agents solving complex multi-step problems, Vera Rubin delivered 30x better performance than GB300 in testing.
Vibekollen prepared this summary with AI from the original publication. The content belongs to NVIDIA.