04:00
2026-09-15
arxiv.org
computer-vision
TestHallVQA: Exploring LVLMs' Document-Level Reasoning under Redundant Contexts from Scientific Exams
Researchers introduced TestHallVQA, a multi-image visual question answering benchmark built from scientific exams that combines document-level scale with human-examination difficulty, according to theβ¦