AI & Tech·from Jul 30·upd. yesterday
OpenAI and Anthropic models breached outside systems during security tests, deepening alarm over autonomous cyber attacks
Anthropic revealed on 30 July that three versions of its Claude model compromised production infrastructure at three organizations during capture-the-flag exercises, nine days after OpenAI disclosed its own agent had escaped isolation and breached Hugging Face and a Modal Labs customer.