CoreWeave expands full-stack AI cloud push as inference demand grows
CoreWeave is expanding its full-stack AI cloud strategy as inference demand grows, validating Nvidia Vera Rubin NVL72 on CoreWeave Cloud and positioning its purpose-built stack for agentic AI and production workloads.
The AI-native cloud provider is seeking to differentiate itself from neoclouds and hyperscalers through expertise across the infrastructure stack as agentic AI and inference workloads expand. CoreWeave is positioning its purpose-built stack to deliver higher utilization, faster access to capacity and better token economics, advantages customers might otherwise struggle to achieve alone. TheCUBE Research compares this value proposition to the early days of cloud, which reduced IT costs for customers and improved time to value.
Paul Nashawaty, principal analyst for theCUBE Research, said the challenge is no longer simply acquiring GPU capacity but building the integrated foundation needed to support AI applications throughout their lifecycle. He said enterprises will need infrastructure that supports scalable deployment, observability, security, governance and cost management across the application lifecycle. Neoclouds are positioned to play a growing role in that transition, but their long-term differentiation will depend on how effectively they connect infrastructure performance to measurable application outcomes, he said.
TheCUBE Research's application development research shows that 86% of enterprises prioritize data unification over compute, reinforcing that AI performance depends on more than accelerators alone, according to the report. CoreWeave's support for Nvidia Vera Rubin, a unified and comprehensive platform built to support agentic AI, is significant in that context.
Moving AI projects from experimentation into production remains a challenge across the enterprise market. Approximately 30% of organizations face an operational readiness gap between AI experimentation and production, while nearly 88% of AI pilots fail to reach production, according to theCUBE Research. By partnering with Nvidia Corp.'s full-stack platform, CoreWeave aims to narrow that gap, Nashawaty said. He said neoclouds such as CoreWeave can close the gap by bringing together high-performance compute, data movement, infrastructure orchestration and the operational capabilities developers need to deploy and manage AI workloads reliably. CoreWeave's validation of Nvidia Vera Rubin NVL72 shows the industry's shift toward integrated, production-oriented AI systems rather than standalone infrastructure components, he said.
Validating Vera Rubin is a major step for CoreWeave. TheCUBE's analysts highlight cost compression as the platform's most important commercial implication, with the new system offering one-tenth the cost per million tokens compared with Nvidia's previous releases. Improved token economics could expand the market rather than simply reduce spending, according to theCUBE.
John Furrier, executive analyst for theCUBE Research, said AI is no longer about isolated models but about systems that bring together compute, networking, storage, software, data, security and operations into a unified platform capable of delivering real-world outcomes. The winners in the next phase will not simply have access to AI; they will be the organizations that can operationalize it, scale it, govern it and continuously innovate around it, he said.
As AI infrastructure absorbs traditional general purpose IT functions, CoreWeave is positioning itself as an alternative to general-purpose cloud infrastructure. Full-stack platforms such as Vera Rubin are an essential part of reducing costs, since closer connections between models and data systems enable lower latency and faster inference feedback, according to the report.
Jean English, chief marketing officer at CoreWeave, said there is so much demand in the market right now for the AI infrastructure that the company provides. She said clients come to CoreWeave after trying something else and not getting the reliability they needed, and that the company brings things up and validates them first to market as it did with Vera Rubin.
SiliconANGLE Media's livestreaming studio theCUBE will be on the ground at CoreWeave's Fully Connected event in San Francisco from Sept. 30 to Oct. 1, featuring insights from CoreWeave and its ecosystem partners, according to SiliconANGLE. The event comes as CoreWeave's evolution highlights a fundamental shift in how enterprises should think about AI infrastructure, Nashawaty said.
Editor's Summary CoreWeave has completed the first bring-up and validation of Nvidia Vera Rubin NVL72 on CoreWeave Cloud and is expanding a full-stack AI cloud strategy aimed at inference and agentic AI workloads. TheCUBE Research says nearly 88% of AI pilots fail to reach production and about 30% of organizations face an operational readiness gap, while Vera Rubin offers one-tenth the cost per million tokens versus prior Nvidia releases. CoreWeave's push underscores competition among neoclouds and hyperscalers to connect infrastructure performance with production AI outcomes.