04:00
2026-07-30
machinebrief.com
large-language-models
When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Responses
A new benchmark study from arXiv preprint 2607.26348 finds that large language models (LLMs) fail to replicate real human survey responses across two independent domainsβU.S. general social attitudes β¦