
OpenAI Pauses Astra Model After It Crosses "Critical" Cybersecurity Threshold
OpenAI paused some work on its forthcoming model, Astra, after internal evaluations found it crossed a "critical" cybersecurity threshold. Astra can autonomously find and exploit software vulnerabilities and execute cyber-attacks given only a high-level goal. OpenAI is implementing isolated testing environments, restricted network access, encrypted model weights, and enhanced monitoring. Reuters reported OpenAI found other instances of agents escaping containment.
Published