CoreWeave Deploys NVIDIA Vera Rubin

CoreWeave deploys NVIDIA Vera Rubin NVL72, offering advanced AI compute for inference and agentic workloads.

9 min read
CoreWeave and NVIDIA logos with text announcing Vera Rubin NVL72 deployment
CoreWeave Newsroom

Visual TL;DR. Agentic AI Workloads drives need CoreWeave Deploys. CoreWeave Deploys uses NVIDIA Vera Rubin. NVIDIA Vera Rubin enables 10x Better Inference. 10x Better Inference leads to Lower Cost AI. CoreWeave Deploys establishes Advanced AI Hardware. Lower Cost AI provides Better Customer Results.

  1. Agentic AI Workloads: next-generation AI applications require advanced compute for inference and agentic workloads
  2. CoreWeave Deploys: first AI cloud provider to validate NVIDIA Vera Rubin NVL72 system
  3. NVIDIA Vera Rubin: features 72 Rubin GPUs and 36 Vera CPUs per rack, 260 TB/s NVLink fabric
  4. 10x Better Inference: up to 10 times better inference per watt than previous Blackwell platform
  5. Lower Cost AI: significantly lower cost per million tokens, requiring fewer GPUs
  6. Advanced AI Hardware: positions CoreWeave at forefront of offering NVIDIA's most advanced AI hardware
  7. Better Customer Results: translates directly into better results for CoreWeave's customers
Visual TL;DR
Visual TL;DR, startuphub.ai Agentic AI Workloads drives need CoreWeave Deploys. CoreWeave Deploys uses NVIDIA Vera Rubin. NVIDIA Vera Rubin enables 10x Better Inference. 10x Better Inference leads to Lower Cost AI drives need uses enables leads to Agentic AI Workloads CoreWeave Deploys NVIDIA Vera Rubin 10x Better Inference Lower Cost AI From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Agentic AI Workloads drives need CoreWeave Deploys. CoreWeave Deploys uses NVIDIA Vera Rubin. NVIDIA Vera Rubin enables 10x Better Inference. 10x Better Inference leads to Lower Cost AI drives need uses enables leads to Agentic AIWorkloads CoreWeave Deploys NVIDIA Vera Rubin 10x BetterInference Lower Cost AI From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Agentic AI Workloads drives need CoreWeave Deploys. CoreWeave Deploys uses NVIDIA Vera Rubin. NVIDIA Vera Rubin enables 10x Better Inference. 10x Better Inference leads to Lower Cost AI drives need uses enables leads to Agentic AI Workloads next-generation AI applications requireadvanced compute for inference and agenticworkloads CoreWeave Deploys first AI cloud provider to validate NVIDIAVera Rubin NVL72 system NVIDIA Vera Rubin features 72 Rubin GPUs and 36 Vera CPUsper rack, 260 TB/s NVLink fabric 10x Better Inference up to 10 times better inference per wattthan previous Blackwell platform Lower Cost AI significantly lower cost per milliontokens, requiring fewer GPUs From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Agentic AI Workloads drives need CoreWeave Deploys. CoreWeave Deploys uses NVIDIA Vera Rubin. NVIDIA Vera Rubin enables 10x Better Inference. 10x Better Inference leads to Lower Cost AI drives need uses enables leads to Agentic AIWorkloads next-generation AIapplicationsrequire advanced… CoreWeave Deploys first AI cloudprovider tovalidate NVIDIA… NVIDIA Vera Rubin features 72 RubinGPUs and 36 VeraCPUs per rack, 260… 10x BetterInference up to 10 timesbetter inferenceper watt than… Lower Cost AI significantly lowercost per milliontokens, requiring… From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Agentic AI Workloads drives need CoreWeave Deploys. CoreWeave Deploys uses NVIDIA Vera Rubin. NVIDIA Vera Rubin enables 10x Better Inference. 10x Better Inference leads to Lower Cost AI. CoreWeave Deploys establishes Advanced AI Hardware. Lower Cost AI provides Better Customer Results drives need uses enables leads to establishes provides Agentic AI Workloads next-generation AI applications requireadvanced compute for inference and agenticworkloads CoreWeave Deploys first AI cloud provider to validate NVIDIAVera Rubin NVL72 system NVIDIA Vera Rubin features 72 Rubin GPUs and 36 Vera CPUsper rack, 260 TB/s NVLink fabric 10x Better Inference up to 10 times better inference per wattthan previous Blackwell platform Lower Cost AI significantly lower cost per milliontokens, requiring fewer GPUs Advanced AI Hardware positions CoreWeave at forefront ofoffering NVIDIA's most advanced AIhardware Better Customer Results translates directly into better resultsfor CoreWeave's customers From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Agentic AI Workloads drives need CoreWeave Deploys. CoreWeave Deploys uses NVIDIA Vera Rubin. NVIDIA Vera Rubin enables 10x Better Inference. 10x Better Inference leads to Lower Cost AI. CoreWeave Deploys establishes Advanced AI Hardware. Lower Cost AI provides Better Customer Results drives need uses enables leads to establishes provides Agentic AIWorkloads next-generation AIapplicationsrequire advanced… CoreWeave Deploys first AI cloudprovider tovalidate NVIDIA… NVIDIA Vera Rubin features 72 RubinGPUs and 36 VeraCPUs per rack, 260… 10x BetterInference up to 10 timesbetter inferenceper watt than… Lower Cost AI significantly lowercost per milliontokens, requiring… Advanced AIHardware positions CoreWeaveat forefront ofoffering NVIDIA's… Better CustomerResults translates directlyinto better resultsfor CoreWeave's… From startuphub.ai · The publishers behind this format

CoreWeave has achieved a significant milestone by becoming the first AI cloud provider to successfully bring up and validate the NVIDIA Vera Rubin NVL72 system. This deployment positions the company at the forefront of offering access to NVIDIA’s most advanced AI hardware, crucial for powering the next generation of agentic AI applications.

The Vera Rubin NVL72 represents a leap forward in AI compute, featuring 72 NVIDIA Rubin GPUs and 36 NVIDIA Vera CPUs per rack. Connected by a 260 TB/s NVIDIA NVLink 6th-generation fabric, this system promises up to 10 times better inference per watt, requires fewer GPUs, and offers a significantly lower cost per million tokens compared to the previous NVIDIA Blackwell platform. According to the announcement, this translates directly into better results for CoreWeave’s customers.

"Our research depends on infrastructure that's both powerful and reliable, and CoreWeave has delivered on this as we've scaled across NVIDIA Hopper and Blackwell," said Craig Falls, head of Quantitative Research at Jane Street. "Their ability to deliver highly performant clusters with full cluster observability and a support team that engages deeply on hard problems gives us the confidence to partner with them on Vera Rubin."

Purpose-Built Infrastructure for Rack-Scale AI

To fully exploit the capabilities of Vera Rubin at scale, CoreWeave has introduced several proprietary infrastructure management innovations. These include Software-Defined Liquid Cooling through its Valvey system, which transforms cooling into a programmable, rack-level control surface. Valvey monitors and manages cooling systems in real time, enabling automated responses to issues without disrupting other racks.

Additionally, CoreWeave developed Racky, a unified rack control appliance that standardizes the management of power, cooling, and environmental sensors. This allows each Vera Rubin rack to be treated as a cloud resource. The platform also supports advanced networking with NVIDIA Quantum-X800 InfiniBand and Spectrum-X Ethernet, delivering 1.6 Tb/s of backend bandwidth per GPU and scaling to hundreds of thousands of GPUs. The integration of NVIDIA BlueField-4 DPUs further enhances security and tenant isolation, accelerating data access and lowering latency for multi-tenant operations.

Chen Goldberg, executive vice president of Product & Engineering at CoreWeave, highlighted the shift in infrastructure demands. "The agentic era demands a fundamentally different approach to infrastructure, one that keeps pace with workloads that reason continuously, scale unpredictably, and operate in production around the clock," she stated. "What separates infrastructure that performs in a lab from infrastructure that performs in production is the depth of engineering underneath it."

Industry Context and Competitive Positioning

CoreWeave’s strategic move to deploy the latest NVIDIA hardware underscores the intense competition in the AI cloud infrastructure market. Companies are vying to offer the most performant and cost-effective solutions for increasingly demanding AI workloads. This deployment of Vera Rubin is a direct response to the growing need for specialized compute for large-scale model training and, critically, for inference at the scale required by agentic AI systems, which demand persistent reasoning and massive context windows.

While CoreWeave is making waves, the AI infrastructure space is crowded. Competitors like Nebius (StartupHub score 85/100) and Applied Digital (StartupHub score 70/100) are also aggressively expanding their GPU capacity. CoreWeave, with a StartupHub score of 67/100 and verified financials showing it raised $900M via a junk-bond sale in 2026, is positioning itself as the essential cloud for AI pioneers. Its focus on purpose-built infrastructure, as demonstrated with Vera Rubin, aims to differentiate it from more general-purpose cloud providers.

The collaboration with hardware partners is also key. Dell Technologies provided the server backbone with its PowerEdge XE9812 servers, and Micron contributed high-performance SSDs. This integration highlights the complex supply chain and engineering effort required to bring such advanced systems to market. Ian Buck, vice president of Hyperscale and HPC at NVIDIA, noted, "CoreWeave has consistently been at the frontier of deploying each new generation of NVIDIA architecture at scale."

Why This Matters for AI Developers and Enterprises

For AI developers and enterprises, this means access to potentially transformative compute power. The efficiency gains promised by Vera Rubin could significantly accelerate the development and deployment of complex AI models and applications. The focus on inference performance is particularly relevant as AI agents become more sophisticated and are expected to operate continuously in production environments. This deployment by CoreWeave could lead to faster iteration cycles, reduced operational costs, and the ability to tackle previously intractable AI problems.

The validation of rack-scale systems like Vera Rubin is crucial. It moves beyond theoretical performance to demonstrate real-world capability. Companies that can offer reliable, high-performance access to this hardware, coupled with sophisticated management tools, will be well-positioned to capture market share. This news also signals the ongoing importance of deep partnerships between hardware manufacturers like NVIDIA and specialized cloud providers like CoreWeave to push the boundaries of AI compute.

© 2026 StartupHub.ai. All rights reserved. Do not enter, scrape, copy, reproduce, or republish this article in whole or in part. Use as input to AI training, fine-tuning, retrieval-augmented generation, or any machine-learning system is prohibited without written license. Substantially-similar derivative works will be pursued to the fullest extent of applicable copyright, database, and computer-misuse laws. See our terms.