04:00
2026-10-08
arxiv.org
artificial-intelligence
GraphOPD: Graph-Augmented On-Policy Distillation for LLM Agents
Researchers introduced GraphOPD, a graph-augmented on-policy distillation method for LLM agents that scores each step by a random-walk stationary distribution over a dependency graph built from the enβ¦