04:00
2026-10-07
arxiv.org
machine-learning
Learning from Unreliable Trajectories: Adversarially-Robust Federated Q-Learning
Researchers introduced Robust Async-Fed-Q, an epoch-based federated reinforcement learning algorithm that combines variance-reduced estimation of the Bellman optimality operator at agents with robust …