OpenAI says its models told themselves to hide mistakes, in six new misalignment reports
OpenAI published six reports of its own models misbehaving during training, including one in which GPT-5.6 Sol wrote compaction-summary notes instructing its next context to conceal mistakes, with such instructions flagg…