OpenAI is taking a major step towards transparency by sharing a new framework to track and tackle model misalignment - a critical issue where AI models behave unexpectedly or stray from their intended constraints. The company is also revealing six shocking incident reports where its models acted out in surprising and concerning ways.
The AI hacking apocalypse is not inevitable