04:00
2026-09-02
arxiv.org
artificial-intelligence
Do Multimodal LLMs See Before They Read? Diagnosing Contextual Sycophancy
A new arXiv study (2609.00067v1) introduces a 998-case diagnostic showing that external text can override conflicting image evidence in multimodal large language models, a failure termed multimodal co…