14:32
2026-07-14
lesswrong.com
ai-safety
Synthetic Scalable Oversight
Researchers at Stagira Labs propose synthetic scalable oversight, a technique that creates graphical abstractions of real-world problems to train tiny models as proxies for evaluating oversight protocβ¦