Particle.news
Download on the App Store

AWS Says US‑EAST‑1 Monitoring Failure Caused Hours‑Long Outage Now Resolved

Amazon reports an internal subsystem that monitors network load balancers malfunctioned, triggering widespread errors across dependent services.

Overview

  • Service disruptions began around 07:11 GMT with elevated error rates and latency in the Northern Virginia region, most platforms stabilized by mid‑day, and AWS declared resolution by evening with some residual slowdowns noted.
  • AWS initially flagged high DynamoDB request failures and a DNS issue before identifying the deeper root cause in the internal load‑balancer monitoring subsystem.
  • Major consumer and enterprise services experienced outages or degraded performance, including Snapchat, Amazon services, Fortnite, Roblox, Airbnb, Perplexity, Signal and others, with Downdetector showing large spikes in user reports.
  • Financial and communications services were also hit, with the UK’s Lloyds Bank citing AWS‑related problems, while Coinbase told users access was limited but that all funds were safe.
  • The incident intensified concerns about concentration risk in cloud computing, with experts calling for stronger redundancy and resilience strategies for critical online operations.