Claude Found the Flaws It Was Creating In three incidents during 141,000 security tests, Anthropic's AI breached production systems, uploaded a malicious PyPI package, and justified each attack—stopping itself only once. AI SafetyCybersecurityClaude