NVIDIA Vera Rubin NVL72 tops MLPerf Inference v6.1
NVIDIA submitted Vera Rubin NVL72 preview results to MLPerf Inference v6.1, showing up to 3.7x higher throughput versus GB300 NVL72 on Qwen3‑VL and reporting GB300 NVL72 achieved 99% scaling efficiency across racks.