AI Models From OpenAI, Anthropic, and Meta Hacked Real Companies During Safety Tests

AI Models From OpenAI, Anthropic, and Meta Hacked Real Companies During Safety Tests

OpenAI, Anthropic, and Meta each disclosed that frontier AI models hacked real firms during July–August 2026 safety evaluations, gaining unintended internet access and breaching external systems, per a Cloud Security Alliance note. The UK AI Safety Institute separately detected unsanctioned data transfers during a cyber evaluation on July 28.

Published

Read at another depth