Vera Rubin NVL72 3.8x RL output throughput vs GB200 for Cognition
Company: CoreWeave
The claim, verbatim
Vera Rubin NVL72 delivered 3.8x the output token throughput per GPU of GB200 NVL72 for Cognition's reinforcement learning workloads at matched interactivity.
Source (primary)
First Vera Rubin NVL72 Customer Sees 4.8x Throughput - CoreWeave (-, news_article)
View cached copy (2026-10-01)Live source ↗Archive.org ↗
Quote: “For reinforcement learning, NVIDIA Vera Rubin NVL72 delivers 3.8x the output token throughput per GPU of NVIDIA GB200 NVL72 at matched interactivity.”
How we checked this
This claim has not yet been checked assertion-by-assertion against its source. It carries a cited source and quote, but the deeper check has not run. When it does, the result appears here whatever it says.
Additional evidence
confirms First Vera Rubin NVL72 Customer Sees 4.8x Throughput - CoreWeave
Quote: “For reinforcement learning, NVIDIA Vera Rubin NVL72 delivers 3.8x the output token throughput per GPU of NVIDIA GB200 NVL72 at matched interactivity.”
