惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

V
V2EX
P
Proofpoint News Feed
D
DataBreaches.Net
C
Check Point Blog
L
LangChain Blog
量子位
美团技术团队
Vercel News
Vercel News
人人都是产品经理
人人都是产品经理
N
Netflix TechBlog - Medium
V
Visual Studio Blog
Microsoft Security Blog
Microsoft Security Blog
博客园 - 【当耐特】
MongoDB | Blog
MongoDB | Blog
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
Last Week in AI
Last Week in AI
The GitHub Blog
The GitHub Blog
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
U
Unit 42
腾讯CDC
M
MIT News - Artificial intelligence
Microsoft Azure Blog
Microsoft Azure Blog
Blog — PlanetScale
Blog — PlanetScale

AI demand is so high, AWS customers are trying to buy out its entire capacity | Network World

Cisco: Latest news and insights 2026 network outage report and internet health check Selector targets the network visibility gap in multi-cloud infrastructure AI reshapes cybersecurity workforce priorities as IT teams brace for new risks Top network and data center events of 2026 How AI is transforming network incident response (and where it still falls short) Google opens TPUs to enterprises beyond its own cloud via Blackstone JV AI, cybersecurity skills top IT pay premiums Startup Bolt Graphics promises 5x performance over Nvidia’s best GPU Wireless security is a battle of AI vs. AI NetOps teams look to AI to automate Day 2 operations Digital twins reshape network and data center management Network outages, power failures strain data center resiliency Five takeaways from Cisco's blowout quarter and what it means to customers Cisco to cut nearly 4,000 jobs despite strong growth in AI, enterprise networking Startup SPAN teams with Nvidia to put data center nodes in your backyard Hard drive shortage affecting enterprise storage needs Wi-Fi 8 is closer than you think. Here’s what you need to know Cisco open-sources agentic AI security spec HPE revamps private cloud stack for enterprises rethinking VMware Versa takes aim at fragmented enterprise security with CSPM, orchestration update, and AI agent controls Red Hat opens Ansible to AI agents, within limits Red Hat offers endless Linux support — for a fee Red Hat: Sovereignty is more than just compliance Tech job postings hit three-year high as AI demand fuels hiring rebound HPE memory server targets compute-heavy and agentic AI workloads PCI group begins work on new spec to support bandwidth-hungry apps like AI, HPC Q&A: Quantum physicist Sonia Fernández-Vidal on why classical computing isn't going anywhere OpenAI-led consortium seeks to address AI processing bottlenecks Gluware's Titan rises to meet Mythos network vulnerability challenge
AWS hit by US-East-1 outage after data center thermal event
2026-05-08 · via AI demand is so high, AWS customers are trying to buy out its entire capacity | Network World

A power outage triggered by a thermal event inside an Amazon Web Services data center in Northern Virginia disrupted Elastic Compute Cloud (EC2) instances and Elastic Block Store (EBS) volumes in the US-EAST-1 region late on Thursday, the cloud provider confirmed in updates posted to its Health Dashboard.

In an incident report timestamped 5:25 PM PDT (00:25 UTC Friday), AWS said it had spotted issues in the use1-az4 availability zone and confirmed that “EC2 instances and EBS volumes hosted on impacted hardware are affected by the loss of power during the thermal event.” Rising temperatures inside a single data center had caused the impairments, the company said in a statement.

AWS shifted traffic away from the affected zone for most services and warned of longer-than-usual provisioning times.

As the evening progressed, the company struggled to bring temperatures down. By 6:47 PM PDT, AWS warned that “Other AWS services that depend on the affected EC2 instances and EBS volumes in this Availability Zone may also experience impairments,” and at 8:06 PM PDT, it conceded that “progress is slower than originally anticipated,” recommending that customers needing immediate recovery restore from EBS snapshots or launch resources in unaffected zones.

By 10:11 PM PDT, AWS reported “incremental progress to restore cooling systems” but said users were still “experiencing elevated error rates and latencies for some workflows.”

The May 7 incident is not the first time US-EAST-1 has gone down. The region suffered two outages in October 2025, including a 15-hour disruption on October 19 and 20 caused by a race condition in DynamoDB’s automated DNS management system that affected over 70 AWS services and produced cascading failures across Slack, Atlassian, Snapchat, and other dependent services. AWS regions in Ohio have also experienced power-related outages tied to EC2 instances in past years.

Customer services go dark

As recovery progressed through the night, AWS confirmed that some services were coming back online faster than others.

“Some AWS services, such as IoT Core, ELB, NAT Gateway, and Redshift, continue to see significant improvements in the recovery of their workflows,” AWS said in a later update. “However, some customers will continue to see their affected EC2 instances and EBS volumes as impaired until we achieve full recovery.”

KoboToolbox, a data collection platform used by humanitarian and development organisations, said its global instance went offline at 00:32 UTC on May 8 because of the AWS infrastructure problem, according to a community advisory posted by Kobo staff. The platform’s EU instance was unaffected.

Physical-layer risk gets a fresh look

Such outages are not unique to AWS, said Bhuvie Chhabra, senior principal analyst at Gartner. “All major cloud providers have experienced similar incidents, highlighting the inherent complexity and challenges of operating at hyperscale,” Chhabra said.

The May 7 event raises a question CISOs should not assume away. CISOs should assess “to what degree AZs are located in physically distinct facilities versus coexisting within the same physical data center” and whether each zone has independent power, networking, cooling, and physical security, Chhabra said. Even when virtual instances are redundant across zones, an application will fail if its database is not similarly redundant, he added.

Kaustubh K, practice director at Everest Group, said physical-layer failures should push enterprises to broaden their resilience playbooks. “Physical-layer failures such as power and cooling disruptions highlight that enterprises should extend resilience planning beyond software and cyber risks, particularly for mission-critical applications,” he said. CISOs should identify critical workloads where infrastructure-level disruptions could materially impact operations and ensure appropriate redundancy, failover, and recovery mechanisms are built into the architecture, Kaustubh added.

Concentration risk back in focus

What sets US-EAST-1 apart from other AWS regions is the weight of global dependencies it carries. Many AWS global services, including Identity and Access Management authentication, CloudFront, Route 53, and DynamoDB Global Tables, depend on US-EAST-1 endpoints even for resources deployed in other regions, AWS confirmed in updates during the October 2025 incident.

US-EAST-1 is a critical global dependency for AWS, and except for Oracle, all hyperscale providers carry some global dependencies, Chhabra said. AWS is unique in publicly documenting these in its Fault Isolation Boundaries white paper. “Reducing the concentration risk to zero is unattainable,” Chhabra said, adding that CISOs must instead manage it through a life cycle approach to third-party risk management, partnering with sourcing, procurement, and vendor management to track changes in the vendor footprint.

“While Availability Zone separation continues to provide an important resilience layer, enterprises running mission-critical workloads should periodically reassess regional concentration risk and validate whether their resilience posture aligns with business continuity expectations,” Kaustubh said.

SUBSCRIBE TO OUR NEWSLETTER

From our editors straight to your inbox

Get started by entering your email address below.