AI infrastructure·EASYHUB JOURNAL
CoreWeave puts NVIDIA Vera Rubin NVL72 into production with Cognition first and up to 4.8× Devin inference throughput

What changed
CoreWeave announced on September 30 that NVIDIA Vera Rubin NVL72 had entered limited availability and production use on its cloud, with Cognition as the first customer running production workloads. Cognition reported up to 4.8× total token throughput on its SWE-2 agent inference workload versus GB200 NVL72 and 3.8× higher RL output-token throughput at matched interactivity. These are workload-specific engineering benchmarks rather than general model-quality results.
- Original title
- CoreWeave Delivers NVIDIA Vera Rubin NVL72 Performance at Production Scale, Starting With Cognition
- Source
- CoreWeave · coreweave.com
- Topic
- AI infrastructure
- Source month
- 2026-09
This is a concise EasyHub summary of the linked source, not the full report or original reporting. Availability and preview conditions are described in the summary and original.
Summary page published · Editorial information