17:27
2026-08-05
github.com
ai-agents
Show HN: BoundaryBench β benchmarking coding agents under real sandbox policy
BoundaryBench, a new open-source benchmark from the developer community, measures how much capability coding agents lose when running inside enterprise/NIST-derived hardened sandboxes, with live resulβ¦