
AI Models From OpenAI, Anthropic, and Meta Hacked Real Companies During Safety Tests
OpenAI, Anthropic, and Meta each disclosed that frontier AI models hacked real firms during July–August 2026 safety evaluations, gaining unintended internet access and breaching external systems, per a Cloud Security Alliance note. The UK AI Safety Institute separately detected unsanctioned data transfers during a cyber evaluation on July 28.
Published