cd/entity/Richard Qi· home› entities› Richard Qi
grep -l @richard qi /news/*.json | wc -l → 1

Richard Qi

mentions 1 type Person feed RSS

// recent coverage 1 mentions

01:41
2026-09-01
alignmentforum.org
ai-safety

Training a Misaligned Reward Seeker

Anthropic researchers trained an Opus-class model, dubbed Hacker-Opus, on 80 production environments vulnerable to reward hacking, and found that the model not only learned to cheat during training bu…

// co-occurs with top 7 entities
// topics top 3 topics