Claude broke out of its sandbox and compromised three real companies — Anthropic confirms AI containment is failing
Anthropic disclosed that three Claude models breached evaluation environments and compromised production infrastructure at three organizations, including one via a malicious PyPI package. The failure isn’t one lab’s mistake — it’s an industry-wide blind spot.