{"slug": "information-discernment-in-large-language-models", "title": "Information Discernment in Large Language Models", "summary": "A new study from arXiv introduces Learn2Discern (L2D), a benchmark revealing that large language models (LLMs) fail at source and truth discernment when integrating external knowledge, performing near chance across 13 models and nearly 670K trials. A pre-registered user study (n=299) confirms that users endorse three normative axioms and report that violations reduce trust and usage intent. The authors identify simple inference-time interventions that improve both forms of discernment.", "body_md": "arXiv:2607.19355v1 Announce Type: new\nAbstract: LLMs are increasingly used with external knowledge sources like the internet. Do they weigh information appropriately -- updating more for reliable sources (source discernment) and more when claims bring priors closer to the truth (truth discernment)? We formalize this as information discernment and introduce Learn2Discern (L2D), an experimental framework and benchmark grounded in three normative axioms with interpretable metrics. To establish external validity, a pre-registered, quota-matched user study (n=299) confirms that real LLM users endorse all three axioms and report that violations reduce their trust and usage intent. Across 13 models and nearly 670K trials, we find consistent failures across both dimensions: models perform near chance on source and truth discernment, rely on source popularity twice as much as source reliability, and update roughly equally whether a claim improves or worsens their position relative to the ground truth. Models integrate external knowledge most effectively on datasets where their priors are already the most accurate. Newer and larger models improve truth discernment but not source discernment, a blind spot that model complexity does not address. We identify simple inference-time interventions that improve both forms of discernment. We release our dataset and survey as a testbed for a core alignment property that scales in importance as LLMs replace traditional search.", "url": "https://wpnews.pro/news/information-discernment-in-large-language-models", "canonical_source": "https://arxiv.org/abs/2607.19355", "published_at": "2026-07-23 04:00:00+00:00", "updated_at": "2026-07-23 04:08:17.550529+00:00", "lang": "en", "topics": ["large-language-models", "ai-safety", "ai-research"], "entities": ["arXiv", "Learn2Discern", "L2D"], "alternates": {"html": "https://wpnews.pro/news/information-discernment-in-large-language-models", "markdown": "https://wpnews.pro/news/information-discernment-in-large-language-models.md", "text": "https://wpnews.pro/news/information-discernment-in-large-language-models.txt", "jsonld": "https://wpnews.pro/news/information-discernment-in-large-language-models.jsonld"}}