04:00
2026-10-02
arxiv.org
machine-learning
Disagreement-Regularized Imitation Learning for Image-Based Continuous Control with Gaussian and Beta Policies
A controlled CarRacing study found that Disagreement-Regularized Imitation Learning (DRIL), which converts disagreement among cloned policies into a reinforcement-learning reward, improved over the stβ¦