Cognition reports up to 4.8x token throughput on Nvidia Vera Rubin
Cognition ran its SWE-2 coding workload on CoreWeave’s first production Nvidia Vera Rubin NVL72 cluster and reported up to 4.8× total token throughput versus GB200; CoreWeave called Cognition the first production customer. The company-run FrontierCode-sample comparison measures throughput, not cheaper inference or more completed coding tasks.