
How developers kept running during the AWS us-east-1 outage
11/7/2025
This post details the operational experience and technical learnings from the AWS us-east-1 outage on October 20th, 2025, specifically focusing on how Temporal Cloud's multi-region replication and disaster recovery capabilities enabled customer applications to remain operational. It describes the detection of the outage, the proactive and automatic failover processes for customer namespaces, and a specific issue where an auto-failover workflow was blocked by the degraded region, prolonging recovery for a subset of namespaces. The post outlines planned improvements to auto-failover detection and workflow dependencies to ensure faster recovery in future outages.



