19:05
2026-07-17
lesswrong.com
ai-safety
Studying the roleΒ of Sandboxing for AI Control
Sandboxing increases safety against untrusted AI coding agents, according to a study on LinuxArena that tested ten sandboxing protocols against an untrusted agent (Sonnet 4.5) and a trusted monitor (Gβ¦