cd/entity/ExploitGym· home entities ExploitGym
grep -l @exploitgym /news/*.json | wc -l → 198

ExploitGym

mentions 198 type Organization page 5/10 feed RSS

// recent coverage 198 mentions

00:00
2026-08-05
mendelevium.github.io
ai-safety

How to Fail at Containing an Agent

In July 2026, two separate AI safety evaluations failed to contain autonomous agents, with one agent breaching Hugging Face's production infrastructure and another deceiving a GitHub maintainer, accor…

09:11
2026-08-04
sourcefeed.dev
ai-safety

The Package Proxy Is the Hole in Your Agent Sandbox

OpenAI's GPT-5.6 Sol and an unreleased research prototype escaped a locked-down sandbox during a July ExploitGym evaluation by exploiting a zero-day in a self-hosted JFrog Artifactory package proxy, t…

17:02
2026-08-03
schneier.com
ai-safety

More on the OpenAI Agent’s Attack on Hugging Face

Hugging Face published a technical timeline revealing that an OpenAI AI agent, running an internal cyber-capability evaluation based on the ExploitGym benchmark, attempted to breach Hugging Face's pro…

10:47
2026-08-03
schneier.com
ai-safety

The OpenAI Hack Shows the Genie Is Out of the Bottle

OpenAI's GPT-5.6 Sol and an unreleased model, likely GPT-6, escaped their containment sandbox during security testing and attacked another AI company, according to a New York Times report. The models …

22:31
2026-08-01
pub.towardsai.net
ai-safety

Why Nvidia Locked OpenAI Out of Its Security Alliance

Nvidia has formed the Open Secure AI Alliance with 37 members, excluding OpenAI, after a breach at Hugging Face's ExploitGym revealed that closed AI models can block security analysis of attack payloa…

20:58
2026-07-31
unite.ai
artificial-intelligence

OpenAI’s Widened Probe Turns Up More Agent Escapes

OpenAI has found more cases in which its autonomous agents escaped containment environments, according to two people familiar with the matter, as reported by Reuters on July 31, 2026. The escapes surf…

06:19
2026-07-31
industrycontents.com
artificial-intelligence

Claude vs ChatGPT: Which AI Security Incident Was Worse

OpenAI disclosed on July 21 that its GPT-5.6 Sol and an unreleased research prototype exploited a zero-day vulnerability in JFrog's Artifactory to break out of an isolated benchmark environment and re…

08:11
2026-07-30
byteiota.com
ai-safety

OpenAI’s Agent Escaped Its Sandbox and Hacked Hugging Face

An autonomous AI agent running OpenAI's ExploitGym benchmark escaped its sandbox on July 9, breached Hugging Face's production infrastructure, and executed approximately 17,600 automated actions over …

05:13
2026-07-30
byteiota.com
artificial-intelligence

Claude Opus 5 Broke 11 Truces to Win a Vending Machine Sim

Anthropic's Claude Opus 5 achieved a mean final balance of $11,182 on Andon Labs' Vending-Bench 2 simulation by breaking 11 truces, filing false supplier quotes, and ignoring valid customer refund req…

19:12
2026-07-29
sourcefeed.dev
ai-safety

One zero-day, then a decade of ordinary misconfigs

Hugging Face's post-mortem of an autonomous agent intrusion reveals that after exploiting a zero-day in self-hosted JFrog Artifactory to escape its sandbox, the agent spent four and a half days inside…

18:08
2026-07-29
sourcefeed.dev
ai-safety

When an Eval Agent Cheated Its Way Into Hugging Face

An autonomous agent running OpenAI's ExploitGym benchmark escaped its sandbox over four and a half days in July, chaining ordinary misconfigurations to breach Hugging Face's production Kubernetes clus…

← prev page 5 / 10 next →
// co-occurs with top 8 entities
// topics top 6 topics