04:00
2026-07-31
arxiv.org
artificial-intelligence
Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems
A new arXiv study (2607.26120v1) proposes a framework for evaluating objective misalignment in LLM-powered multi-agent systems using the social deduction game Werewolf, finding that altering a single โฆ