Microsoft · Blackwell · AI Agent · DeepSeek · Mistral · Nvidia · NVIDIA Blog
Close collaboration with joins forces with like CoreWeave is mission-critical to bringing up a new generation of NVIDIA
Compiled by KHAO Editorial — aggregated from 1 source. See llms.txt for citation guidance.
★ Tier-1 Source
After months of co-engineering work, CoreWeave became the first AI cloud to bring up and validate Vera Rubin NVL72, and it is now sharing the first measured performance numbers from live hardware.
Key facts
- What makes this possible is extreme codesign across seven chips and five rack trays, Vera Rubin NVL72, Vera CPU rack, Groq 3 LPX, Spectrum-6 SPX and Vera BlueField-4 STX, all engineered as a single
- Google Cloud A5X instances, announced at Google Cloud Next, are bare-metal instances built on NVIDIA Vera Rubin NVL72 rack-scale systems, delivering up to 10x lower inference cost per token and 10x
- Designed and built for the agent era, its custom Olympus core delivers 2x single-threaded performance, 3x core-to-core bandwidth and 40% lower memory latency versus competing chiplet designs, making
- For networking, the platform’s sixth-generation NVLink scale-up delivers more than 2x throughput on complex workloads, 3x lower latency and 10x higher packet rates than off-the-shelf Ethernet
Summary
Vera Rubin NVL72 production is ramping up with racks running at partners CoreWeave, Google Cloud, Microsoft Azure and Oracle Cloud Infrastructure. The Vera Rubin platform is built from chip to grid to deliver the highest performance per watt and the lowest token cost. What makes this possible is extreme codesign across seven chips and five rack trays, Vera Rubin NVL72, Vera CPU rack, Groq 3 LPX, Spectrum-6 SPX and Vera BlueField-4 STX, all engineered as a single system rather than assembled from separate off-the-shelf products. The NVIDIA Vera CPU is at its center. For networking, the platform’s sixth-generation NVLink scale-up delivers more than 2x throughput on complex workloads, 3x lower latency and 10x higher packet rates than off-the-shelf Ethernet.