Higher inference throughput
Leading tokens/sec per GPU for DeepSeek-R1 in server and offline mode on NVIDIA GB300 NVL72, among all v6.0 submitters
Record-setting performance
Largest-ever submission with 8,192 NVIDIA GB300 NVL72 GPUs, reaching the MLPerf quality target for the DeepSeek-V3 671B benchmark in 2.02 minutes
2.80x faster training
On the same Llama 3.1 405B benchmark, CoreWeave improved training time from 27.33 minutes in v5.0 to 9.77 minutes in v6.0, with 1.70x more useful work per GPU-minute






















