Skip to content
VibekollenBETAVibekollen
BlogNVIDIA

NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut

Article image or reusable cover for NVIDIA

NVIDIA presents its new Vera Rubin NVL72 system, which shows significant performance improvements in MLPerf Inference v6.1 tests.

The system delivers up to 3.7x higher throughput than its predecessor GB300 NVL72 on the Qwen3-VL benchmark, and 2.5x higher on DeepSeek-R1. When four GB300 NVL72 racks were scaled up together, 99 percent scaling efficiency was achieved — throughput grew nearly linearly as more GPUs were added. NVIDIA also achieved performance gains through software optimization: version 6.1 was up to 1.6x faster than 6.0. For AI agents solving complex multi-step problems, Vera Rubin delivered 30x better performance than GB300 in testing.

Read the full story at NVIDIA →

Vibekollen prepared this summary with AI from the original publication. The content belongs to NVIDIA.

More to read