04:00
2026-09-22
arxiv.org
large-language-models
Recognition, Simulation, and Refusal: A Contamination-Aware Study of Classic Psychological Effects in LLM Agents
A new benchmark called PsyAgentBench, presented in arXiv paper 2609.22090v1, re-ran classic psychology experiments on LLM agents across 41,904 trials and found that apparently human-like effects ariseβ¦