04:00
2026-08-28
arxiv.org
machine-learning
Shared Actors Need Not Share Critics: Effects of Value Mismatch in Parallel Reinforcement Learning
A new arXiv paper (2608.26481v1) finds that using a single shared critic across parallel environments in reinforcement learning causes value mismatch that degrades learning, and shows that conditionin…