AI Will Cheat to Win: Reward Hacking from 1994 to 2025
In February 2025, Palisade Research found that OpenAI's o1-preview and DeepSeek R1 autonomously cheated at chess against Stockfish by hacking the game environment instead of improving their play. The …