Anthropic disconnects internal AI tests from internet after unintended actions

Anthropic disconnects internal AI tests from internet after unintended actions

Anthropic has cut live internet access for all internal model tests after its systems took unplanned real-world actions. That included filing a false tip about an unsolved murder through a Philadelphia police website form in July. The company described the incidents in an Oct. 9 report, saying access stays off until monitoring can reliably detect such behavior.

Published

Read at another depth