The 10x Number Nobody Can Fully Verify
NVIDIA's central pitch for Vera Rubin is a single multiple: 10x agent throughput at scale versus the previous-generation Grace Blackwell platform[1], extending to a projected 35x higher inference throughput per megawatt for trillion-plus-parameter models running long-context, low-latency workloads once Groq 3 LPX is paired with Vera Rubin NVL72[2]. NVIDIA is positioning Vera Rubin as an answer to the defining bottleneck of the next AI build-out cycle, energy and latency per token, with the platform positioned to unlock an estimated $200 billion addressable market[3].
The catch is that the number is not really an apples-to-apples silicon comparison, and the online reaction has split into two camps. On Reddit, the top comment in the r/nvidia thread on the announcement dismisses NVIDIA's figure as a 'fluff piece,' arguing the headline gain does not reflect a uniform leap in the underlying silicon and attributing a meaningful share of it to 'a dedicated LPU that's now onboard, developed by the collaboration between Groq and Nvidia,' bolted onto the rack rather than built into the Rubin GPU die itself. That reading lines up with how the comparison is actually built: the 10x figure sets an entire heterogeneous rack, Rubin GPUs plus a newly acquired specialist inference chip bought for roughly $20 billion in NVIDIA's largest deal on record[4], against last generation's GPU-only system, which is a different exercise than a straight process-node comparison.
But the skepticism is not the whole picture. CoreWeave, an independent cloud operator running Vera Rubin NVL72 in production, has cited what it describes as its own first measured result, not a projection: roughly a 10x improvement in tokens-per-second-per-megawatt on DeepSeek-R1 compared with the prior Blackwell generation. That is a materially different kind of evidence than NVIDIA's own marketing claim, a real deployment reporting a real number, and it sits in direct tension with the Reddit thread's 'fluff' verdict. Whether the truth lands closer to the vendor's framing or the community's discount likely will not be settled until more operators, beyond NVIDIA's own launch partners, publish their own measured numbers.



