04:00
2026-09-30
arxiv.org
computer-vision
LeRF: Learning Reference Coordinate Frames for Perspective Taking Reasoning
A new arXiv paper (arXiv:2609.36219v1) introduces LeRF, a framework that trains Vision-Language Models to build and use explicit reference coordinate frames for perspective-taking reasoning, addressin…