
7/23/2026
What this post added
This post provides the first measured silicon performance statistics for the NVIDIA Vera Rubin NVL72 on CoreWeave, demonstrating a 10x improvement in tokens per megawatt for the DeepSeek R1 inference workload compared to NVIDIA GB200 NVL72. It details the architectural advantages of Vera Rubin NVL72, including its GPU count, CPU count, NVLink fabric bandwidth, and NVFP4 support, and highlights the impact of these on inference performance and cost-efficiency for agentic AI workloads. The post also notes that this is an initial measurement and further optimizations are expected.