Issue 2026-09-16 · Industry · 芯片 · 研究 · 产品

NVIDIA Vera Rubin NVL72 tops MLPerf Inference v6.1

NVIDIA Vera Rubin NVL72 tops MLPerf Inference v6.1

NVIDIA submitted Vera Rubin NVL72 preview results to MLPerf Inference v6.1, showing up to 3.7x higher throughput versus GB300 NVL72 on Qwen3‑VL and reporting GB300 NVL72 achieved 99% scaling efficiency across racks.

NVIDIA Blog12 h ago
Read original ↗