04:00
2026-08-14
arxiv.org
artificial-intelligence
Agreement Is Not Alignment: Divergent Moral Grounds in Human and LLM Ethical Judgments
A new arXiv paper (2608.12368v1) finds that high agreement between large language models and human annotators on ethical judgments does not imply alignment, as models systematically diverge in the morβ¦