04:00
2026-09-11
arxiv.org
computer-vision
BodyCam-VQA: Enhanced Body-Worn Camera Video Captioning via Multimodal Reasoning and Probe Question Generation
A new arXiv paper (2609.10815v1) proposes BodyCam-VQA, an adaptive visual question answering framework designed to extract fine-grained forensic evidence from police body-worn camera footage that curr…