Drooid Logo
Back to story perspectives

Full Breakdown

Amazon Web Services Faces Scrutiny Over AI-Related Outages

2/20/2026, 10:09:35 PM

Overview of the Outages

Amazon Web Services (AWS) experienced at least two significant outages in December 2025, attributed to errors involving its internal AI tools. The most notable incident lasted 13 hours and occurred when engineers permitted the Kiro AI coding tool to autonomously make changes, resulting in the deletion and recreation of an environment. This disruption primarily affected a system used by customers to explore AWS service costs and was reported by the Financial Times.

Company Response and Justification

AWS has firmly attributed these outages to user error, specifically misconfigured access controls, rather than faults in the AI technology itself. An AWS spokesperson stated that the incidents were "extremely limited" and did not impact core services such as compute, storage, or databases. The company emphasized that the Kiro tool, which is designed to assist developers, typically requests authorization before executing actions. However, in these cases, the permissions granted to engineers allowed for actions without the usual oversight.

Internal Concerns and Expert Opinions

Despite AWS's assertions, internal sources within the company have expressed concerns regarding the autonomy granted to AI tools like Kiro. A senior AWS employee noted that the outages were "entirely foreseeable" and highlighted the risks associated with allowing AI to resolve issues without human intervention. Experts have also raised doubts about the reliability of AI systems in critical operational contexts. Cybersecurity expert Michal Wozniak pointed out that while human engineers can recognize potential errors during manual processes, AI lacks the contextual awareness necessary to prevent significant mistakes.

Broader Implications for Cloud Infrastructure

These incidents have sparked discussions about the operational risks and governance challenges that arise as cloud providers increasingly integrate AI-driven tools into their infrastructure. The reliance on automation for managing complex systems necessitates robust controls and oversight to mitigate the potential for errors. AWS has reportedly implemented additional safeguards, including mandatory peer reviews and staff training, following the December outages.

Criticism of AI Integration

Critics argue that Amazon's framing of the layoffs and the role of AI in its operations raises questions about the company's commitment to human oversight. While AWS has stated that the outages were coincidental and not indicative of broader issues with AI, skepticism remains among employees and industry experts regarding the effectiveness and safety of AI tools in production environments.

Verbatim Quotes

  • “The company: “This brief event was the result of user error—specifically misconfigured access controls—not AI.” — AWS Spokesperson
  • “But when a slop generator is involved in an outage, suddenly that’s just ‘coincidence’,” he added.” — Michal Wozniak, Cybersecurity Expert

Conclusion

The recent outages at AWS underscore the complexities and risks associated with the integration of AI tools in cloud infrastructure management. As the company continues to expand its use of automation, the need for stringent oversight and governance becomes increasingly critical to ensure reliability and maintain customer trust.