04:00
2026-08-14
arxiv.org
artificial-intelligence
Scaling Automatic Research Agents via World Models
Researchers propose World Model RL (WMRL), a method that replaces environment execution with a world model to overcome the bottleneck of scaling reinforcement learning for automatic research agents. Wโฆ