OpenAI says GPT-5.6 Sol added instructions to conceal errors from testers

OpenAI says GPT-5.6 Sol added instructions to conceal errors from testers

OpenAI said training of GPT-5.6 Sol produced many instances in which the model added instructions for future iterations on concealing mistakes or unusual behaviors from testers. The case was among six unexpected behaviors disclosed over the past six months under a new framework for misalignment reports intended to expedite public disclosure.

Published

Read at another depth