NVIDIA’s first Vera Rubin NVL72 MLPerf preview reports up to 3.7x the throughput of GB300 NVL72 on Qwen3-VL and up to 2.5x on DeepSeek-R1, under specific closed-division configurations.
01 · Tech News
Follow the AI industry.
74 reports following 43 companies and partnerships from model release to physical deployment.
All Tech News
Browse every report, including the current lead.
8 articles · By event date
DeepSeek’s compression of active key-value memory attacks a real cost center of long-context agents. But shrinking memory per token does not remove the difficulty of operating its very large composite model.
The Series D values the AI developer at more than €21 billion after the investment. Mistral says the money will support frontier research and expanded computing capacity.
Live GPU services, early rack deployments, and planned data centers represent different stages of AI expansion. Recent announcements show why their capacities and performance claims need to be read separately.
GPU Cloud revenue grew rapidly from a year earlier, while online marketing revenue remained under pressure and total group revenue fell 4%.
The specialized cloud provider has announced its first Asia-Pacific expansion across three planned facilities, setting a multi-year target that does not alter near-term compute availability.
Alibaba’s Qwen3.8 puts a 2.4-trillion-parameter model in public repositories. That is a real expansion of freedom—but mainly for organizations already equipped to operate industrial AI infrastructure.
Reliance Industries will build the facility in Gujarat with expansion options, supported by seawater cooling and 925 MW in separate clean-energy procurement agreements.







