04:00
2026-08-18
arxiv.org
artificial-intelligence
When to Communicate: Belief Distributions and KL Divergence for Principled Gating in Multi-Agent RL
A new arXiv preprint (2608.14559v1) proposes a principled gating mechanism for multi-agent reinforcement learning where agents communicate only when the KL divergence between their learned belief distβ¦