Vera Rubin NVL72 4.8x token throughput (Cognition test)
Company: CoreWeave
The claim, verbatim
Cognition's testing on CoreWeave showed Vera Rubin NVL72 delivered up to 4.8x higher total token throughput for SWE-2 inference versus GB200 NVL72, and 3.8x higher output-token throughput for RL workloads.
Source (primary)
CRWV's Vera Rubin Push: Can AI Infrastructure Fuel Its Growth Engine? - Yahoo Finance (-, news_article)
View cached copy (2026-10-04)Live source ↗Archive.org ↗
Quote: “the Vera Rubin NVL72 delivered up to 4.8x higher total token throughput for SWE-2 inference workloads versus a GB200 NVL72 baseline, 3.8x higher output-token throughput for reinforcement-learning workloads”
How we checked this
This claim has not yet been checked assertion-by-assertion against its source. It carries a cited source and quote, but the deeper check has not run. When it does, the result appears here whatever it says.
Additional evidence
confirms CRWV's Vera Rubin Push: Can AI Infrastructure Fuel Its Growth Engine? - Yahoo Finance
Quote: “the Vera Rubin NVL72 delivered up to 4.8x higher total token throughput for SWE-2 inference workloads versus a GB200 NVL72 baseline, 3.8x higher output-token throughput for reinforcement-learning workloads”
