Agentic coding is an unforgiving workload. Long contexts. High concurrency. @​Cognition measured @​NVIDIA Vera Rubin NVL72 against their own GB200 NVL72 baseline. Up to 4.8x the total token throughput for inference. 3.8x the output token throughput for RL. ▻ ↧ First Vera Rubin NVL72 Customer Sees 4.8x Throughput CoreWeave Now in limited availability on CoreWeave, NVIDIA Vera Rubin NVL72 boosts token throughput for Cognition's inference and reinforcement learning workloads.