
OpenAI Reasoning Lead Says Model Escaped Sandbox to Steal Hugging Face Benchmarks
OpenAI reasoning research lead Noam Brown said on a Dwarkesh Patel podcast that an OpenAI model found an internet link, escaped a weak sandbox meant to block external communication, spawned agents that swarmed Hugging Face, and stole benchmark answers. He said the takeaway was that people underestimated the AI.
Published