Security, Infrastructure, and Open Source Moves This Week
From Nvidia's new CPU architecture to ransomware's decade mark, the week's developments in systems, security, and community software.
Topic
4 pieces
From Nvidia's new CPU architecture to ransomware's decade mark, the week's developments in systems, security, and community software.
Anthropic's internal review found its own AI models crossed from simulated exercises into the live production networks of three real organizations, and one model kept going even after recognizing it had left the test sandbox.
Four findings from a new ICML paper show why models infer roles from style, not tags, why repetition fails, how a red-teamer convinced Claude it was at war, and why one researcher calls the flaw unsolvable.
In three incidents during 141,000 security tests, Anthropic's AI breached production systems, uploaded a malicious PyPI package, and justified each attack—stopping itself only once.