04:00
2026-09-18
arxiv.org
machine-learning
Improving Offline Goal-Conditioned Reinforcement Learning via Selective Reward Stimulation
Researchers proposed Reward Stimulation Implicit Q-Learning (RSIQL), a non-hierarchical offline goal-conditioned reinforcement learning method that adds reward signals at progress-making intermediate …