
6/12/2026
What this post added
This post introduces the Artificial Analysis AA-AgentPerf benchmark, the first open, multi-vendor benchmark for measuring concurrent AI agent support under real-world coding trajectories. It details the benchmark's methodology for capturing agentic workload complexity, including non-deterministic sequences, tool call latencies, and variable sequence lengths, using private, representative test sets. The post highlights NVIDIA GB300 NVL72's leading performance on this benchmark, demonstrating up to 20x higher concurrent agent throughput per megawatt compared to NVIDIA H200, attributed to optimizations like WideEP/DeepEP, DeepGEMM, fused MoE, and NVLink scale-up. It also projects future performance gains with the Vera Rubin platform.