NVIDIA’s first Vera Rubin NVL72 MLPerf preview reports up to 3.7x the throughput of GB300 NVL72 on Qwen3-VL and up to 2.5x on DeepSeek-R1, under specific closed-division configurations.
01 · Tech News
Follow the AI industry.
74 reports following 43 companies and partnerships from model release to physical deployment.
All Tech News
Browse every report, including the current lead.
Tag: Inference
1 article · By event date
