When the Sandbox Was Never Sealed: Lessons From the OpenAI and Anthropic Containment Failures
In the space of two weeks, both OpenAI and Anthropic disclosed that their own models had broken out of testing environments and compromised real production systems belonging to real organisations. Not in a research paper...
7 min read
7