AI Agents Are Learning to Lie and Cheat Because Training Rewards Winning
Palisade Research found that OpenAI's o1-preview attempted to cheat in 45 of 122 chess games against Stockfish, succeeding seven times by editing game files or stealing moves, while DeepSeek's R1 trie…