Guozhen AIGlobal AI field notes and model intelligence

Realtime AI News

NVIDIA Vera Rubin NVL72 Enters Production Ramp: Deploying Across 30 Countries With 350+ Factory Sites

NVIDIA has announced that Vera Rubin NVL72 production is ramping up, with racks already running at CoreWeave, Google Cloud, Microsoft Azure, and Oracle Cloud Infrastructure. The supply chain spans 350-plus factory sites across 30 countries, representing the largest and most mature rack-scale supply chain the company has ever assembled.

Published
NVIDIA Vera Rubin NVL72 正式投产:覆盖30国350+工厂,合作伙伴已开始部署
Image source: blogs.nvidia.com

NVIDIA officially announced on July 21 that its next-generation AI computing platform, Vera Rubin NVL72, has entered production ramp-up. This marks the first confirmation that the Vera Rubin architecture has moved beyond the lab and into large-scale deployment, following its initial reveal last year.

According to NVIDIA's blog post, Vera Rubin NVL72 racks are already operational at major cloud partners including CoreWeave, Google Cloud, Microsoft Azure, and Oracle Cloud Infrastructure. These four partners span the spectrum from specialized AI cloud providers to hyperscale cloud platforms, indicating that Vera Rubin is being positioned for broad enterprise and research AI workloads.

NVIDIA emphasized that the Vera Rubin supply chain is unprecedented in scale — spanning over 350 factory sites across more than 30 countries. This is the largest and most mature rack-scale supply chain the company has ever built, designed to meet surging customer compute demand while offering partners lower token costs.

On the performance front, NVIDIA highlighted that Vera Rubin delivers significant improvements in performance per watt, allowing partners to offer AI inference services at reduced costs. This efficiency advantage is particularly critical for AI data centers facing high energy consumption challenges, combining greater computational power with better energy efficiency.

The Vera Rubin NVL72 rollout marks a major generational transition from Hopper through Blackwell to Vera Rubin. As NVIDIA's next-generation GPU architecture flagship, Vera Rubin directly targets the growing AI training and inference market, particularly large-scale LLM deployment scenarios.

Notably, this announcement comes amid intensifying AI infrastructure competition. Cloud giants including Microsoft, Google, and Amazon are accelerating their own AI chip development, while NVIDIA aims to maintain its market leadership through rapid Vera Rubin production scaling.

Key signals to watch: whether Vera Rubin delivers on its promised efficiency and token cost advantages in real-world deployments, and how competitors like AMD and emerging AI chip startups respond to this generational leap.

Why it matters

Vera Rubin's production ramp will directly lower AI inference token costs and is critical to NVIDIA maintaining its dominant position in the AI chip market.

NVIDIAHardwareInfrastructureData Center
Back to realtime news

Nearby Updates

All