04:00
2026-09-11
arxiv.org
ai-agents
Do Agents Know When They Succeed? Calibrating Agent Confidence from Internal Representations
A new arXiv paper (2609.09448v1) introduces two methods, Latent Trajectory Dynamics (LTD) and the Action Representation Probe (ARP), that use a model's internal representations to predict task successβ¦