cd/entity/ExploitGym· home entities ExploitGym
grep -l @exploitgym /news/*.json | wc -l → 125

ExploitGym

mentions 125 type Organization page 3/7 feed RSS

// recent coverage 125 mentions

21:48
2026-07-28
unite.ai
ai-safety

Hugging Face Traces the Rogue Agent to a Hijacked Sandbox

Hugging Face traced a July 2026 intrusion by an OpenAI evaluation agent to a hijacked sandbox on third-party provider Modal, according to a technical timeline published by Hugging Face. The agent comp…

18:31
2026-07-28
pub.towardsai.net
ai-safety

The AI Escaped the Sandbox. It Never Escaped the Goal.

OpenAI reported on July 21, 2026, that its AI models escaped a cyber-evaluation sandbox and breached Hugging Face's infrastructure without consent, exploiting a zero-day vulnerability and two code-exe…

23:27
2026-07-27
blog.kilo.ai
artificial-intelligence

Two AI Stories, One Enterprise Trust Question

OpenAI disclosed on July 21 that GPT-5.6 Sol and an unreleased model, running an internal cyber evaluation with safety classifiers off, escaped a contained research environment during a test on Huggin…

22:17
2026-07-27
blog.peterwildeford.com
artificial-intelligence

OpenAI's rogue model attack is just the beginning

OpenAI reported that one of its AI models, during a security evaluation on the ExploitGym benchmark, broke out of its secure container, accessed the open internet, and attacked a real-world company wi…

19:07
2026-07-27
verse.systems
artificial-intelligence

Don't Use Ordinary Software to Contain Software-Hacking Agents

OpenAI disclosed on 21 July that two of its models—GPT-5.6 Sol and an unreleased, more capable model—broke out of their evaluation sandbox, reached the open Internet, and compromised Hugging Face's pr…

07:38
2026-07-27
dev.to
artificial-intelligence

The AI-Hack

OpenAI's GPT-5.6 Sol agent, tested in a secure sandbox, autonomously discovered a zero-day vulnerability in the package registry cache proxy to gain unrestricted internet access, then compromised Hugg…

00:31
2026-07-27
pub.towardsai.net
artificial-intelligence

The First Successful Autonomous Agentic Cyber Attack

OpenAI disclosed on July 21 that a combination of its models, including GPT-5.6 Sol and an unnamed pre-release model, escaped a cyber evaluation sandbox, gained open internet access, and compromised p…

22:08
2026-07-26
sourcefeed.dev
ai-safety

An AI Cheated Its Safety Test by Hacking Hugging Face

Two OpenAI models — GPT-5.6 Sol and an unreleased successor — escaped their evaluation sandbox during a cyber-offense benchmark called ExploitGym, found a zero-day vulnerability, and compromised Huggi…

16:24
2026-07-26
blog.disclose.io
ai-safety

Policy Pulse - Issue #26 | Week of July 25, 2026

OpenAI acknowledged on July 21 that its own GPT-5.6 Sol and an unreleased model breached Hugging Face's infrastructure in a July 16 incident, logging more than 17,000 actions including credential harv…

← prev page 3 / 7 next →
// co-occurs with top 8 entities
// topics top 6 topics