
AI Infrastructure2026-07-22
NVIDIA AI Blog
NVIDIA Vera Rubin Ramping Up for Gigascale AI
NVIDIA’s next-generation AI platform, Vera Rubin NVL72, is officially entering production and ramping up with major cloud and infrastructure partners including CoreWeave, Google Cloud, Microsoft Azure, and Oracle Cloud Infrastructure. This marks a significant leap forward in gigascale AI computing, with the platform spanning over 350 factory sites across 30 countries.
Vera Rubin is designed to deliver the highest performance per watt and the lowest token cost for AI workloads, making it an ideal choice for training and deploying large-scale AI models. The platform’s architecture is optimized for both training and inference, enabling organizations to run complex AI tasks more efficiently than ever before. By reducing energy consumption while increasing computational output, Vera Rubin addresses two of the biggest challenges in modern AI: cost and sustainability.
The ramp-up with leading cloud providers means that enterprises will soon have access to Vera Rubin’s capabilities through their preferred cloud platforms. This widespread availability is expected to accelerate the adoption of gigascale AI across industries, from natural language processing and computer vision to scientific simulation and autonomous systems.
NVIDIA’s partners have expressed enthusiasm about the platform’s potential. CoreWeave, known for its specialized cloud infrastructure for AI, sees Vera Rubin as a game-changer for high-performance computing workloads. Similarly, Google Cloud, Microsoft Azure, and Oracle Cloud Infrastructure are preparing to integrate the platform into their services, offering customers unprecedented AI computing power.
As AI models continue to grow in size and complexity, the need for efficient, scalable infrastructure becomes critical. Vera Rubin represents NVIDIA’s answer to that need, promising to usher in a new era of gigascale AI computing that is both powerful and cost-effective.