NVIDIA Vera Rubin Driving Performance Per Watt, Lowest Token Cost for Partners Worldwide

NVIDIA Vera Rubin Driving Performance Per Watt, Lowest Token Cost for Partners Worldwide 图片 1
NVIDIA Vera Rubin Driving Performance Per Watt, Lowest Token Cost for Partners Worldwide 图片 2
NVIDIA Vera Rubin Driving Performance Per Watt, Lowest Token Cost for Partners Worldwide 图片 3
NVIDIA Vera Rubin Driving Performance Per Watt, Lowest Token Cost for Partners Worldwide 图片 4

NVIDIA Vera Rubin is here, and it’s going gigascale.

Vera Rubin NVL72 production is ramping up with racks running at partners CoreWeave, Google Cloud, Microsoft Azure and Oracle Cloud Infrastructure. Spanning 350+ factory sites in 30 countries, Vera Rubin has the largest, most mature rack-scale supply chain ever assembled to meet customer compute demand.

The Vera Rubin platform is built from chip to grid to deliver the highest performance per watt and the lowest token cost. CoreWeave’s first benchmark on DeepSeek-R1 says it all: 10x more throughput per megawatt than Grace Blackwell NVL72 — landing directly on the metric that matters most for power-constrained AI factories.

Advancing Performance With Extreme Codesign

What makes this possible is extreme codesign across seven chips and five rack trays — Vera Rubin NVL72, Vera CPU rack, Groq 3 LPX, Spectrum-6 SPX and Vera BlueField-4 STX — all engineered as a single system rather than assembled from separate off-the-shelf products.

The NVIDIA Vera CPU is at its center. It redefines what an AI factory CPU can be. Designed and built for the agent era, its custom Olympus core delivers 2x single-threaded performance, 3x core-to-core bandwidth and 40% lower memory latency versus competing chiplet designs, making it the most efficient single-threaded CPU for the agentic workloads that matter most.

Accelerating AI Factories With Purpose-Built Networking

For networking, the platform’s sixth-generation NVLink scale-up delivers more than 2x throughput on complex workloads, 3x lower latency and 10x higher packet rates than off-the-shelf Ethernet. For scale-out, Spectrum-X Ethernet combines 102.4T Spectrum-6 switch systems, 1.6T ConnectX-9 SuperNICs, adaptive routing, advanced congestion control, telemetry and open operating system support, enabling 1.6x higher RDMA bandwidth than off-the-shelf Ethernet.

The world’s leading AI infrastructure builders — including CoreWeave, Microsoft, SpaceXAI and Tesla — are among the firs…

添加评论
点赞收藏
点踩分享查看原文
评论
?
参与讨论