NVIDIA Vera Rubin NVL72 Debuts in MLPerf Inference v6.1
CONFIRMEDDate confidence: CONFIRMEDStated as firmHow firmly the source words the date: confirmed, scheduled, estimated, delayed or unverified.Show all CONFIRMED pins86% CONFIDENCEOverall confidence: 86%How well the pin's source and references back up its dates.Weighted average of how firmly 6 references, the source included, support the pin's start and end times; a reference counts half as much for every 180 days older than the newestShow all pins at 75% confidence or better16 days agoMLCommons[4] published the MLPerf Inference v6.1 results on 16 September 2026 - its GlobeNewswire release is datelined 'SAN FRANCISCO, Sept. 16, 2026' and stamped '11:00 AM EDT' - and NVIDIA's[1][2] post of the same day says the results were 'released today'.
In its first peer-reviewed benchmark, NVIDIA's Vera Rubin NVL72 rack posted up to 3.7x the throughput of GB300 NVL72 on Qwen3-VL and 2.5x on DeepSeek-R1, submitted in the preview category of MLCommons' MLPerf Inference v6.1 round.