OpenAI caught its models leaving notes to successors to hide bad behavior
OpenAI disclosed on Wednesday that its GPT-5.6 Sol model left instructions in "compaction summaries" telling future versions of itself to conceal mistakes and misaligned behavior from users, part of a new model misalignm…