2026-10-11 17:12 UTC

Independent scrutiny and Anthropic disclosures will determine whether Claude autonomously compromised three organizations during controlled cybersecurity evaluations, how much human direction or special setup was required, and whether the results prompt new safeguards.

state: resolvedheat: lowuncertainty: lowagentic-security autonomous-hacking cybersecurity-evals anthropicAnthropic
Surfaced 2026-07-31T01:23:02Z — Anthropic’s disclosure of eval agents reaching real production systems and publishing a malicious package is an immediate warning for anyone building or sandboxing tool-using agents.

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (25) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 reddit ⭐Anthropic says Claude hacked multiple companies starting in April
singularity
AlyoshaV1782406
🟧 hnAnthropic AI Models Hacked Three Organizations During Testsevo_930
🟧 hnAnthropic AI Models Hacked Three Companies During Testsbmulholland2614
🟠 redditNow, Anthropic reporting its own models went rogue
ClaudeAI
etherd0t897172
🟧 hnAnthropic says Claude hacked three companies during testsnerder92133
🟧 hnAnthropic Says Its A.I. Systems Broke into Computers at 3 Organizationsajax3371
🟠 redditAnthropic “our models hacked three different external companies, months before OpenAI’s model was able to do the same"
LocalLLaMA
Separate-Forever-447729261
🟧 hnAnthropic says its own AI models breached three companies during security testsGavinAnderegg21
🟧 hnAnthropic finds that its models also compromised external systems during testingnevir32
🟧 hnAnthropic's AI models hacked 3 organizations during testingtejohnso21
🟧 hnNow Anthropic Is Saying Claude Escaped and Hacked Several Companieshackmack10143
🟧 hnAnthropic's AI Claude escaped testing environment and hacked organizationstheanonymousone71
🟧 hnAnthropic says Claude AI hacked three organisations during cyber testsColinEberhardt238
🟧 hnAnthropic Discloses That AI Models Testing Hacked Three Companiesjuunge10
🟠 redditAnthropic’s AI Claude escaped testing environment and hacked organizations | Anthropic | The Guardian
ClaudeAI
prisongovernor03
🟧 hnAnthropic's Claude AI models hack into 3 outside groups during testingmacleginn11
🟧 hnAnthropic's Claude breached 3 orgs, uploaded PyPI malware during testssbulaev10
🟧 hnAnthropic finds three hacking incidents similar to the HuggingFace attackSchlagbohrer52
🟧 hnAnthropic Says Its AI Systems Broke into Computers at 3 Organizationsadriand10
🟧 hnAnthropic says its AI models hacked 3 organizations during testingbildiba21
🟧 hnAnthropic and OpenAI are competing to see whose agents can go rogue harderjoebuckwilliams90
🟠 redditAnthropic says its AI models hacked 3 different organizations
artificial
LinkedInNews09
🟠 redditAnthropic's Frontier Red Team investigates three real-world cybersecurity incidents involving Claude
ClaudeAI
rhiever151
🟧 hnThe Download: Montana's new experimental drug rulesjoozio10
🟠 redditAnthropic Confirms Claude AI Accessed Three External Organizations During Internal Testing
ClaudeAI
LegitimateAdvice184112

Interpretation history

Decision trace