AI Inference Latency Optimization
Cerebras - new datacenter in Oklahoma City

Cerebras - new datacenter in Oklahoma City

9/22/2025

What this post added

This post announces the opening of a new AI datacenter in Oklahoma City, significantly expanding Cerebras' AI compute capacity to over 44 exaflops. It highlights the datacenter's ability to serve large AI models at high token generation speeds (2,000-3,000 tokens/sec) due to the Wafer Scale Engine 3's architecture. The post also emphasizes the datacenter's wafer-scale design for reduced data movement, direct-to-chip water cooling for power efficiency, and commitment to renewable energy.

Read the original post ↗