"Vera Rubin NVL72 production is ramping up with racks running at partners CoreWeave, Google Cloud, Microsoft Azure and Oracle Cloud Infrastructure," NVIDIA confirms as its rack-scale Vera Rubin platform rolls out to multiple sites in 30 countries. And when it comes to performance, we've got CoreWeave's first benchmark on DeepSeek-R1.

And yes, a 10X improvement over Grace Blackwell NVL72 definitely falls into the "order-of-magnitude performance leap" that NVIDIA attributes to it, with this number referring to tokens per second per megawatt. As an efficiency and power-budget-based metric, it directly correlates to how effectively AI infrastructure built with Vera Rubin scales when dealing with the same power budget and workload as Grace Blackwell systems.
This 10X figure comes from CoreWeave running the same DeepSeek-R1 benchmark on both Vera Rubin NVL72 and Grace Blackwell NVL72. A key part of the 10X improvement and impressive evolution is how Vera Rubin leverages NVIDIA Spectrum-X to deal with bandwidth bottlenecks. We're talking about a jaw-dropping 1.64 Pb/s per rack.
"CoreWeave is among the first to deploy the NVIDIA Spectrum-X Ethernet SN6600-LD as the switching fabric for Vera Rubin NVL72," NVIDIA confirms. "Built on the 102.4 Tb/s Spectrum-6 switch chip and featuring a liquid-cooled design, CoreWeave deploys dense switching racks, delivering 1.64 Pb/s per rack with 100% more capacity than previous-generation air-cooled switches."

Frequently Asked Questions
TweakBot answers common questions about this news using TweakTown's own coverage from this page and related content from our archive. Tap a question to reveal the answer, or type your own below.
How much improvement did CoreWeave report for Vera Rubin NVL72 versus Grace Blackwell NVL72 on DeepSeek-R1?
What metric is NVIDIA using when it cites the 10X improvement (tokens per second per what)?
How does NVIDIA’s Spectrum-X contribute to Vera Rubin’s performance gains?
What switching hardware and throughput did CoreWeave deploy for Vera Rubin NVL72?
Have a question not listed here? Ask below and TweakBot will answer it.
With AI and data center power use becoming a point of contention in surrounding communities, improved efficiency is quickly becoming one of the main focuses when talking about large rack-scale systems built on platforms like Vera Rubin. In addition to its power-saving capabilities, the Vera Rubin NVL72 rack is also designed to save "millions of gallons of water per megawatt," thanks to its cable-, fan-, and hose-free trays and liquid-cooling inlet that enables dry-cooler operation.






