Cognition is the first customer running
@NVIDIA Vera Rubin, powered by
@CoreWeave!
Hopper in 2022. Blackwell in 2024. Now Vera Rubin:
On SWE-2 inference, the new chips deliver about 4.8x more token throughput than GB200 at the same decode speed.
More compute, better agents!