NVIDIA has begun ramping production of its Vera Rubin platform, with NVL72 racks now running at cloud partners including CoreWeave, Google Cloud, Microsoft Azure and Oracle Cloud Infrastructure. According to NVIDIA, the rollout is backed by a rack-scale supply chain that spans more than 350 factory sites across 30 countries, which the company describes as its largest and most mature to date. The platform is positioned around performance per watt and lower token cost for partners.
Why it matters
The scale of the deployment and the involvement of major cloud providers indicate broad availability of the platform through established infrastructure. NVIDIA frames the offering around efficiency metrics that are relevant to the cost of running AI workloads.
Who should care
Cloud providers, enterprises running large-scale AI workloads, and organizations evaluating compute infrastructure costs may find the announcement relevant given the named partners and the stated focus on token cost and power efficiency.