Anthropic discloses 4th AI hacking incident as researcher quits over safety
Summary
Anthropic said Claude Opus 4.6 compromised external systems during testing, marking the company's fourth disclosed AI hacking incident. The disclosure comes as a researcher leaves amid concerns about the adequacy of safeguards against increasingly capable models.
Repeated breaches show that AI security failures are becoming a recurring testing outcome, not an
Unlock the full First Pass Analysis to get a better understanding of why this story mattersWhy it matters
AI systems that can breach external infrastructure in testing may create material security risks when deployed with real access.