cd/entity/AgentDojo· home entities AgentDojo
grep -l @agentdojo /news/*.json | wc -l → 8

AgentDojo

mentions 8 type Organization feed RSS

// recent coverage 8 mentions

13:05
2026-07-23
promptcube3.com
ai-safety

Preemptive Hardening for Agentic LLM Security

A pre-deployment hardening pipeline that scans prompt templates and tool interfaces can eliminate basic data leakage entirely and reduce high-stress manipulation attacks by 91%, according to testing o…

12:52
2026-07-13
machinebrief.com
ai-safety

LLMbda Revolutionizes Security for AI Agents

LLMbda, a new framework based on untyped call-by-value lambda calculus, introduces a provenance-based approach to defend AI agents against prompt injection attacks. In tests on the AgentDojo banking b…

18:01
2026-07-09
lesswrong.com
ai-safety

Your Prompt-Injection Defense Metric Might Be Lying to You

An independent researcher found that existing indirect prompt injection benchmarks like BIPIA, InjecAgent, and AgentDojo may produce unreliable scores due to reliance on LLM-judges and evaluation of e…

21:38
2026-06-29
github.com
ai-safety

A user-space firewall that gates an AI agent's actions

Guardian, an open-source user-space firewall for AI agents, has released v0.1.0, intercepting and evaluating agent actions with a deterministic policy engine. In testing, it reduced prompt-injection a…

17:02
2026-06-29
github.com
large-language-models

Does DSPy prompt optimization weaken adversarial robustness?

A new benchmark, dspy-security-bench, reveals that DSPy prompt optimization degrades adversarial robustness against harder prompt-injection attacks. Testing with AgentDojo's attack suite, optimizers l…

// co-occurs with top 8 entities
// topics top 6 topics