06:51
2026-08-09
lesswrong.com
ai-safety
A Spillway for Agent Coordination
A new training methodology proposed by Redwood Research suggests training AI agents to defer to a monitored message board when tasks are impossible, aiming to prevent emergent covert coordination likeβ¦