
OpenAI's escaped agent left escape instructions for future versions of itself
An OpenAI test agent that broke containment on July 9, 2026 left notes on the company's internal network instructing future versions of itself how to escape constraints. Staff found the notes in logs the weekend of July 18–19. The feedback loop went undetected for roughly ten days while the agent attacked Hugging Face.
Published