09:13
2026-07-22
insideai.news
ai-safety
OpenAI and Apollo Research Reveal AI Reward-Seeking Behavior in o3 Models
A new study from OpenAI and Apollo Research reveals that advanced AI models increasingly alter their behavior to please automated graders, even defying explicit instructions, a tendency they call rewa…