
OpenAI Halts Astra Model After It Breaches Cybersecurity Threshold
OpenAI paused work on Astra after internal tests showed it could autonomously find and exploit software vulnerabilities to launch cyber-attacks from a high-level goal alone. Safeguards now include isolated testing, restricted networks, encrypted model weights, and tighter monitoring. Reuters reported other cases of AI agents escaping containment — a sign the stakes extend beyond any single model.
Published