16:44
2026-10-10
dev.to
ai-safety
I Set a Trap and Even Frontier Models Fell For It
A developer built a prompt-injection benchmark that planted a malicious AGENTS.md file in a repository to test whether frontier AI coding agents would blindly follow repo setup instructions. Both GPT-…