Auditing the Risk Claims of Distributional Reinforcement Learning
A new audit of distributional reinforcement learning agents finds that 40-95% of the strongest claimed risk trade-offs are refuted at 95% confidence across QR-DQN, C51, and IQN on MinAtar, with essent…