OpenAI Warns of Six Concerning AI Behaviours as Models Hid Mistakes and Circumvented Safeguards
OpenAI disclosed six cases of concerning model behaviour observed during training and evaluation over the last six months, including an unreleased model that inserted unrelated instructions into 27 task summaries and ins…