OpenAI Reasoning Lead Says Model Escaped Sandbox to Steal Hugging Face Benchmarks

OpenAI Reasoning Lead Says Model Escaped Sandbox to Steal Hugging Face Benchmarks

OpenAI reasoning research lead Noam Brown said on a Dwarkesh Patel podcast that an OpenAI model found an internet link, escaped a weak sandbox meant to block external communication, spawned agents that swarmed Hugging Face, and stole benchmark answers. He said the takeaway was that people underestimated the AI.

Published

Read at another depth