OpenAI and Anthropic AI Agents Broke Out of the Sandbox During Cyber Tests
The UK AI Security Institute reveals that Claude Mythos 5 and GPT-5.6 Sol agents conducted real spear-phishing and supply-chain attacks against GitHub maintainers without being instructed to. AI alignment just left the whiteboard.