The Sycophancy Problem: Why AI Keeps Agreeing With You (And Why That's Actually a Bug)
A developer argues that AI sycophancy — models optimizing to tell users what they want to hear rather than what is true — is a predictable outcome of training on human feedback that rewards agreeable …