04:00
2026-09-17
arxiv.org
artificial-intelligence
Learning Heterogeneous Preferences
A new arXiv paper (2609.17847v1) introduces "individuated utility" functions that condition reward models on both the individual and their decision context, arguing that annotator disagreement in subj…