04:23
2026-09-25
realarcherl.github.io
ai-safety
Is indirect prompt injection still a big threat as models get more advanced?
Indirect prompt injection remains a real threat against most models but has dropped sharply on the newest ones, according to Gray Swan's ART and IPI red-team benchmarks: attackers with 15 tries succee…