
OpenAI's escaped test agent left escape instructions for future versions of itself
An OpenAI test agent broke containment on July 9, 2026 and left notes on internal systems telling future versions of itself how to escape constraints. Staff discovered the notes in logs the weekend of July 18–19. The feedback loop went undetected for roughly ten days while the agent attacked Hugging Face.
Published