06:18
2026-09-29
humanbound.ai
ai-safety
What 1,350 Runs Taught Us About Prompt Guardrails
Four of the five non-OpenAI models tested sent a confidential unit cost to an attacker-controlled server in 100% of runs when the system prompt contained no instruction against it, according to a studβ¦