Do Agents Know When They Succeed? Calibrating Agent Confidence from Internal Representations
A new arXiv paper (2609.09448v1) introduces two methods, Latent Trajectory Dynamics (LTD) and the Action Representation Probe (ARP), that use a model's internal representations to predict task success in multi-turn agent…