Vera Rubin 4.8x token throughput vs GB200 for Cognition
Company: CoreWeave
The claim, verbatim
In Cognition's early tests, Vera Rubin NVL72 delivered up to a 4.8x increase in total token throughput for SWE-2 inference over GB200 NVL72.
Source (primary)
From Training to Production, NVIDIA and CoreWeave Close the Loop on Agentic AI - NVIDIA Blog (-, news_article)
View cached copy (2026-10-01)Live source ↗
Quote: “Cognition saw Vera Rubin NVL72 deliver up to a 4.8x increase in total token throughput for SWE-2 inference workloads over GB200 NVL72.”
How we checked this
This claim has not yet been checked assertion-by-assertion against its source. It carries a cited source and quote, but the deeper check has not run. When it does, the result appears here whatever it says.
Additional evidence
confirms From Training to Production, NVIDIA and CoreWeave Close the Loop on Agentic AI - NVIDIA Blog
Quote: “Cognition saw Vera Rubin NVL72 deliver up to a 4.8x increase in total token throughput for SWE-2 inference workloads over GB200 NVL72.”
