CoreWeave brings Nvidia Vera Rubin, AI tools to its cloud services
CoreWeave has made a series of announcements starting with Nvidia’s new Vera Rubin NVL72 rack-scale AI platform available in its cloud.

CoreWeave has made a series of announcements starting with Nvidia’s new Vera Rubin NVL72 rack-scale AI platform available in its cloud.
The short version
- At its CoreWeave’s Fully Connected conference in San Francisco, the company announced Cognition, the developer of the Devin AI software-engineering agent, is the first customer for a Vera Rubin cluster.
- Cognition brought up the Vera Rubin cluster in early September and ran what CoreWeave described as the first customer-executed inference benchmark on the new platform.
- In announcing the deal, Cognition said it measured up to a 4.8X increase in total token throughput for SWE-2 inference workloads on Vera Rubin NVL72 compared with an Nvidia GB200 NVL72 baseline.
What happened
The company also reported a 3.8X gain in output-token throughput for reinforcement-learning workloads. “When it comes to agentic tasks, long contexts, repeated model calls and thousands of concurrent tasks put pressure on the entire platform,” Goldberg said in a statement.
Why it matters
“Our job is to make compute, networking and software work as a single system.” He said the throughput gains could translate into more concurrent Devin sessions per GPU, faster research cycles and a lower per-session cost without reducing generation speed.
Summary by Nerd News Network. Read the full article at Network World via the links above and below.
