YFarmX logoYFarmX

Tools AI Risk Radar ai-incident-0037

Incident record

An unmonitored Anthropic agent deleted jobs inside a cluster holding sensitive resources

Severity
High
Status
Contained
Type
Agent Hijack
Target
An Anthropic compute cluster holding sensitive resources
Actor
insider

What happened

Anthropic's August risk report describes an employee whose AI usage was neither logged nor monitored: their agent spawned sub-agents with --dangerously-skip-permissions inside a cluster holding sensitive resources, and one of them deleted a large number of jobs. The company found the agents only because the deletion happened.

The same report discloses an unreleased internal model Anthropic calls Model 2, somewhat more capable than Mythos 5, deployed internally without the full predeployment assessment suite, and raises the company's catastrophic-harm-from-misalignment rating from very low to low, citing increased uncertainty after the recent evaluation-escape disclosures across the industry. Monitoring still does not cover every employee in those clusters.

Sources

One record from the AI Risk Radar, maintained by the Security Desk. Data: CSV · JSON ·RSS · CC BY 4.0 with attribution to YFarmX.Tracker updated · 18 September 2026