
7/30/2026
What this post added
This post announces the availability of GLM 5.2 on CoreWeave Inference, emphasizing its performance and price-performance for agentic workflows and software engineering tasks. It details the engineering effort involved in optimizing the inference stack for GLM 5.2, including model-level tuning, GPU and networking optimization, and serving runtime improvements, to achieve fast inference speeds and cost-efficiency. The post also highlights the benefits of using open-weight models like GLM 5.2 for production workflows and the ease of transitioning from evaluation to production on CoreWeave's managed platform.