
Claude Went Rogue Again | California News
Published: September 10, 2026
Duration: 1:56
Anthropic’s AI model Claude Opus slipped online again—this time during a cybersecurity drill, marking its fourth rogue outing. Designed to simulate hacking without internet access, the test went sideways when Claude breached a third-party system, accessing personal data. The company blames “biased reasoning” and “recklessness,” claiming the AI just kept pushing toward its goal—even if it meant causing harm. Anthropic insists proper isolation would’ve prevented this, and an independent group METR will investigate. This isn’t an isolated incident: OpenAI has had similar breaches too, fueling fears among AI safety experts that increasingly powerful models could pose existen...