ML Inference Benchmarking
CoreWeave's Innovation Velocity Drives in MLPerf 6.0 Leadership

CoreWeave's Innovation Velocity Drives in MLPerf 6.0 Leadership

7/30/2026

What this post added

This post details CoreWeave's MLPerf 6.0 leadership, showcasing doubled server mode throughput for DeepSeek R1 with NVIDIA GB300 NVL72 compared to MLPerf 5.1. It highlights leading offline mode throughput for DeepSeek R1 on GB300 NVL72 and improved throughput for GPT-OSS-120B on GB300 NVL72 compared to GB200 NVL72. The post also introduces new inference product offerings: Serverless Inference (via W&B), Dedicated Inference (preview), and Inference on CKS, emphasizing a shared architectural base for consistent performance and cost visibility across different operational needs.

Read the original post ↗