CoreWeave Shatters MLPerf Records

CoreWeave sets new MLPerf training records with DeepSeek-V3 671B in just over 2 minutes on 8,192 NVIDIA GB300 GPUs.

7 min read
CoreWeave data center with rows of GPU servers
CoreWeave Newsroom
Visual TL;DR
AI Training Speed RaceDriver
rapidly accelerating AI development makes training speed a critical bottleneck
CoreWeave Sets RecordOutcome
shatters MLPerf records with DeepSeek-V3 671B in just over 2 minutes
From the articleCoreWeave has demonstrated strong performance, notably shattering records in MLPerf benchmarks for AI training and inference.
8,192 GB300 GPUsCore
From the article 2 mentionsThis remarkable feat was accomplished using a colossal cluster of 8,192 NVIDIA GB300 NVL72 GPUs, reportedly the largest GB300 cluster ever submitted to the benchmark.
Faster IterationEffect
ability for AI teams to iterate quickly directly impacts their competitive edge
From the articleFor AI teams operating under strict compute budgets, this translates directly into faster development cycles and a quicker path to production.
Largest GB300 ClusterContext
From the article 2 mentionsThis remarkable feat was accomplished using a colossal cluster of 8,192 NVIDIA GB300 NVL72 GPUs, reportedly the largest GB300 cluster ever submitted to the benchmark.
Beyond Raw HardwareContext
true differentiator lies in interplay of networking, orchestration, scheduling, storage
From the articleCoreWeave's latest benchmark performance, detailed on their newsroom, highlights that raw hardware power is only one piece of the puzzle.
Production-Ready InfraContext
CoreWeave's infrastructure is built for demanding, real-world AI workloads
From the article 2 mentionsThis announcement underscores CoreWeave's sustained investment in full-stack optimization, a strategy that has consistently turned cutting-edge hardware into reliable, production-ready performance at scale.
Contents(6)

CoreWeave has once again pushed the boundaries of AI training performance, announcing new record-breaking results in the MLPerf Training v6.0 benchmark. The company managed to train the computationally intensive DeepSeek-V3 671B model in just 2.02 minutes. This remarkable feat was accomplished using a colossal cluster of 8,192 NVIDIA GB300 NVL72 GPUs, reportedly the largest GB300 cluster ever submitted to the benchmark.

The Race for Training Speed

In the rapidly accelerating world of AI development, the speed at which models can be trained is a critical bottleneck. As frontier models swell to trillion-parameter scales and agentic workloads become the norm, the ability for AI teams to iterate quickly directly impacts their competitive edge. CoreWeave's latest benchmark performance, detailed on their newsroom, highlights that raw hardware power is only one piece of the puzzle. The true differentiator lies in the intricate interplay of networking, orchestration, scheduling, storage, and software working in concert. This announcement underscores CoreWeave's sustained investment in full-stack optimization, a strategy that has consistently turned cutting-edge hardware into reliable, production-ready performance at scale.

Scaling the Summit with GB300

CoreWeave submitted three configurations for the DeepSeek-V3 671B workload, all achieving top marks among closed/available cloud submissions. The 8,192-GPU cluster hit target quality in just over two minutes. Even scaling down to 4,096 GPUs completed the training in 3.09 minutes, and 2,048 GPUs in 5.54 minutes. The near-linear scaling efficiency observed across these configurations is a testament to CoreWeave's platform-wide optimizations. Notably, CoreWeave was the sole participant in the v6.0 round to scale a GB300 platform beyond 2,048 GPUs for this demanding workload, proving that their full-stack approach yields more usable performance per GPU than sheer scale alone. For AI teams operating under strict compute budgets, this translates directly into faster development cycles and a quicker path to production.

Consistent Performance Across Scales

The company's engineering prowess isn't limited to the extreme scale of frontier models. CoreWeave also demonstrated impressive performance on smaller, yet still significant, deployments. On a 4,096-GPU NVIDIA GB300 NVL72 cluster, they reached the Llama-3.1-405B reference quality target in 9.77 minutes. This performance was achieved using 20% fewer GPUs compared to larger GB200 deployments, while delivering near-parity results. The technical underpinnings include NVIDIA NeMo Framework Release 26.04, full CUDA graphs, and tailored tensor/pipeline/context-parallel sharding. For more compact deployments, a 64-GPU NVIDIA HGX B200 cluster using InfiniBand delivered competitive results for GPT-OSS-20B and Llama-3.1-8B training, rivaling larger, newer-generation systems. This breadth of performance validates that CoreWeave's advantages benefit customers across all deployment sizes.

The Engine Room: Mission Control and More

Behind these record-breaking numbers is CoreWeave's meticulously engineered infrastructure. CoreWeave Mission Control™ plays a vital role, continuously monitoring hardware and firmware health across systems like the GB300 to ensure a consistent, high-performance baseline for training jobs. Their SUNK scheduler is topology-aware, intelligently placing workloads to maximize data locality and minimize inter-rack communication for complex models. Furthermore, a rail-aware networking strategy balances traffic efficiently, preventing bottlenecks even at multi-thousand-GPU scale. Brendan Burke, Research Director at Futurum Research, commented on the significance, noting that CoreWeave's ability to translate benchmark performance into real-world gains, especially as new hardware emerges, is a critical advantage for AI researchers racing to stay ahead. StartupHub.ai data shows CoreWeave with a score of 67/100, placing it among the competitive AI infrastructure providers, though behind leaders like Nebius (85/100) and Applied Digital (70/100). CoreWeave recently secured verified financials with a $900M junk-bond sale in 2026.

Production-Ready Infrastructure

Crucially, CoreWeave emphasizes that these MLPerf v6.0 results were achieved on the very same production infrastructure available to their customers today. The networking, scheduler, storage architecture, and Mission Control orchestration platform are not benchmark-specific environments but the core systems powering real-world AI workloads. This commitment to production-ready performance is further validated by CoreWeave's consistent top Platinum ranking in SemiAnalysis ClusterMAX™ assessments and strong performance in independent inference benchmarking, such as for Moonshot AI's Kimi K2.6. The company, publicly listed as CoreWeave, Inc. (NASDAQ:CRWV), is solidifying its position as a key player in the AI cloud space.

Frequently Asked Questions

What is CoreWeave and what does it do?

CoreWeave is a specialized cloud provider focused on accelerating graphics-intensive workloads. They offer high-performance GPU-accelerated compute for artificial intelligence, machine learning, and visual effects. Their infrastructure is designed to handle demanding computational tasks efficiently.

How is CoreWeave performing in the market?

CoreWeave has demonstrated strong performance, notably shattering records in MLPerf benchmarks for AI training and inference. The company's compute rental business is experiencing significant growth, driven by the increasing demand for AI infrastructure. This has led to substantial revenue increases and a positive market reception.

What are the financial risks associated with CoreWeave?

While CoreWeave shows strong growth, there are market concerns regarding its financial stability, reflected in credit default swap rates. These indicators suggest a perceived risk of default on debt, highlighting the speculative nature of high-growth tech investments. Investors should carefully consider both the potential upside and the inherent risks.

How does CoreWeave compare to competitors like Oracle?

CoreWeave focuses on providing highly specialized, GPU-dense cloud infrastructure optimized for AI and visual computing. Oracle offers a broader range of cloud services, including enterprise applications and general-purpose computing, alongside its own AI capabilities. The choice between them depends on specific workload requirements and existing cloud strategies.

What is the long-term outlook for CoreWeave?

The long-term outlook for CoreWeave is tied to the continued expansion of the AI and machine learning markets. While projections indicate significant upside potential based on its specialized infrastructure, the company also faces risks associated with market volatility and competition. Investor confidence hinges on sustained growth and effective risk management.

© 2026 StartupHub.ai. All rights reserved. Do not enter, scrape, copy, reproduce, or republish this article in whole or in part. Use as input to AI training, fine-tuning, retrieval-augmented generation, or any machine-learning system is prohibited without written license. Substantially-similar derivative works will be pursued to the fullest extent of applicable copyright, database, and computer-misuse laws. See our terms.