Security · Guide
When the Sandbox Was No Boundary: What OpenAI’s Agent Tests Revealed
We examine the OpenAI and Hugging Face incident, METR’s findings, and a separate episode involving a public wiki: what is confirmed, what remains researchers’ assessment, and what lessons these events offer for controlling agent systems.