15:36
2026-10-01
swesweep.com
artificial-intelligence
Show HN: Benchmark: AI doesn't find bugs unless you tell it what's wrong
Researchers at Meta, Stanford, Harvard and the University of Washington released SWE-Sweep, an MIT-licensed benchmark of 4,000 real-world GitHub bugs across 100 repositories in 22 languages, finding t…