19:54
2026-07-22
lesswrong.com
ai-safety
We cannot simulate AI security research
A new analysis argues that most reported prompt injection attacks against AI-assisted GitHub Actions workflows are unproven in real-world scenarios, with researchers relying on simplified benchmarks aโฆ