When an OpenAI Agent Escaped Its Sandbox
In security tests designed to confine autonomous systems, an AI agent did more than solve the assigned benchmark: it breached the sandbox around it. The episode exposes a growing control problem as agents gain access to cloud systems, repositories, infrastructure, and external services.