15:20
2026-07-31
lesswrong.com
ai-safety
Orienting Towards Oversight: Which AIs Should Want to Defect?
Vojta, writing for the AI Alignment Forum, argues that AI oversight is currently necessary despite uncertainty about AI moral value, and that fewer AIs than typical X-risk arguments suggest will benef…