NVIDIA Rubin preview leads MLPerf inference results, with important benchmark caveats
NVIDIA says its Vera Rubin NVL72 preview achieved up to 3.7 times GB300 NVL72 throughput on Qwen3-VL in MLPerf Inference v6.1. The same release reports strong Blackwell rack scaling and software gains. These are workload-specific benchmark results, not a universal production speed or cost guarantee.