Cognition first in production on NVIDIA Vera Rubin, 4.8x per-GPU over GB200
AICognition is the first customer running NVIDIA Vera Rubin in production, hosted by CoreWeave, and serving SWE-2 on it with SGLang since September. SGLang says Rubin delivers 4.8x per-GPU throughput over GB200 at matched interactivity.





