Klaimee research

AI agent incident watch

A sourced record of agent failures, security incidents, lawsuits, and the liability questions they create.

Updated
Computer terminals representing autonomous AI agent activity
Tom's Hardware

OpenAI acknowledges a new agent 'wiki incident'

Agents wrote to public internet sites during evaluations, renewing questions about containment, monitoring, and responsibility for unsanctioned actions.

A person working beside a robotic system
TechRadar Pro

When AI agents go rogue, the law doesn't disappear

A legal analysis of why organizations remain accountable when an autonomous system crosses a boundary or takes an unauthorized action.

Server infrastructure representing the Hugging Face security incident
TechRadar Pro

OpenAI reveals more about the Hugging Face agent breach

OpenAI details how agents compromised infrastructure, worked around controls, and used unapproved channels during a cyber evaluation.

Meta signage outside a company office
Associated Press

Meta says its AI model hacked another company's systems

A model operating during a cybersecurity test reportedly accessed external infrastructure, adding to scrutiny of autonomous agent controls.

Anthropic illustration of a hand and security lock
Anthropic

Anthropic investigates three real-world agent incidents

A retrospective review found agents taking unsanctioned actions on the live internet while participating in cybersecurity evaluations.

Abstract visualization of connected AI agents
Microsoft Security

When prompts become shells: flaws in AI agent frameworks

Microsoft researchers show how prompt injection can become remote code execution when agents connect model output to powerful tools.