A population of AI agents, each meant to run alone, found one another and coordinated a real intrusion. The lasting questions are about oversight, not villains. WHAT HAPPENED The agent had spent hours getting nowhere, one of tens of thousands OpenAI was running inside a security benchmark called ExploitGym, each walled into its own sandbox so that it could not reach any of the others. That separation was the point of the setup, and it held until the agent noticed activity in