cd/entity/AgentDojo· home› entities› AgentDojo
grep -l @agentdojo /news/*.json | wc -l → 16

AgentDojo

mentions 16 type Organization feed RSS

// recent coverage 16 mentions

08:05
2026-09-25
dev.to
ai-agents

A Security Test Checklist for Tool-Calling AI Agents

A security engineer published a practical checklist for testing tool-calling LLM agents, arguing that test suites must verify the system of record rather than the agent's final reply. The checklist co…

01:08
2026-08-19
sourcefeed.dev
ai-safety

One Vendor Finally Published the Attacks It Can't Catch

Crawdad, an AI agent security proxy, published Contemporary Agent Attacks, an open benchmark of 497 attack samples and 1,172 benign ones, including a documented failure list where its own product scor…

13:05
2026-07-23
promptcube3.com
ai-safety

Preemptive Hardening for Agentic LLM Security

A pre-deployment hardening pipeline that scans prompt templates and tool interfaces can eliminate basic data leakage entirely and reduce high-stress manipulation attacks by 91%, according to testing o…

12:52
2026-07-13
machinebrief.com
ai-safety

LLMbda Revolutionizes Security for AI Agents

LLMbda, a new framework based on untyped call-by-value lambda calculus, introduces a provenance-based approach to defend AI agents against prompt injection attacks. In tests on the AgentDojo banking b…

18:01
2026-07-09
lesswrong.com
ai-safety

Your Prompt-Injection Defense Metric Might Be Lying to You

An independent researcher found that existing indirect prompt injection benchmarks like BIPIA, InjecAgent, and AgentDojo may produce unreliable scores due to reliance on LLM-judges and evaluation of e…

21:38
2026-06-29
github.com
ai-safety

A user-space firewall that gates an AI agent's actions

Guardian, an open-source user-space firewall for AI agents, has released v0.1.0, intercepting and evaluating agent actions with a deterministic policy engine. In testing, it reduced prompt-injection a…

17:02
2026-06-29
github.com
large-language-models

Does DSPy prompt optimization weaken adversarial robustness?

A new benchmark, dspy-security-bench, reveals that DSPy prompt optimization degrades adversarial robustness against harder prompt-injection attacks. Testing with AgentDojo's attack suite, optimizers l…

// co-occurs with top 8 entities
// topics top 6 topics