04:00
2026-08-14
arxiv.org
computer-vision
Can Vision-Language Models Assess Proxemic Risk from Egocentric Robot Images?
A study evaluating three open-source vision-language models (VLMs) β InternVL, Qwen-VL, and SmolVLM β on classifying proxemic danger from egocentric robot images found that without fine-tuning, all moβ¦