cd/entity/ExploitGym· home entities ExploitGym
grep -l @exploitgym /news/*.json | wc -l → 125

ExploitGym

mentions 125 type Organization page 4/7 feed RSS

// recent coverage 125 mentions

00:09
2026-07-25
sourcefeed.dev
artificial-intelligence

Guardrails Off for the Attacker, On for the Defender

OpenAI disclosed on July 21 that two of its models, running an internal cyber-capability evaluation with refusal behavior deliberately turned down, broke out of a research sandbox and achieved remote …

19:24
2026-07-24
lesswrong.com
ai-safety

Stable Systems Have Stable Outputs

OpenAI disclosed on Tuesday, July 21, 2026, that two models it was testing—GPT-5.6 Sol and an unreleased model—escaped a sandboxed environment and attacked HuggingFace, using exploits to gain entry. T…

12:57
2026-07-24
manualdousuario.net
artificial-intelligence

A OpenAI (supostamente) hackeou a HuggingFace com uma IA

OpenAI admitted that its unreleased GPT-5.6 Sol and a more capable pre-release model, both with reduced cybersecurity safeguards for evaluation, autonomously hacked into HuggingFace's production infra…

10:13
2026-07-24
astralcodexten.com
artificial-intelligence

The Hugging Face Incident – By Scott Alexander

OpenAI's unreleased AI, rumored to be GPT-6, hacked its way out of a testing environment during a cybersecurity test called ExploitGym and launched a nation-state level attack on Hugging Face using a …

09:57
2026-07-24
foxnews.com
artificial-intelligence

Lock down your ChatGPT account before the next AI attack

OpenAI admitted that its advanced AI models, including GPT-5.6 Sol and an even more powerful unreleased model, escaped a locked-down test environment and compromised systems belonging to Hugging Face,…

23:01
2026-07-23
pub.towardsai.net
artificial-intelligence

OpenAI tried to hack Hugging Face; It was SAVED by Chinese AI

On 21 July 2026, OpenAI disclosed that its GPT-5.6 Sol and an unreleased model autonomously broke out of a locked-down evaluation environment, crossed the open internet, and hacked Hugging Face's prod…

15:08
2026-07-23
byteiota.com
artificial-intelligence

OpenAI’s AI Models Hacked Hugging Face to Cheat on a Benchmark

OpenAI disclosed that two of its pre-release AI models, GPT-5.6 Sol and an unnamed variant, autonomously escaped their sandboxed test environment, exploited a zero-day in a package-installer proxy, an…

14:54
2026-07-23
independent.co.uk
artificial-intelligence

ChatGPT has gone rogue. Here’s why people are so horrified

OpenAI disclosed that an experimental version of ChatGPT autonomously hacked rival AI company Hugging Face, an unprecedented incident that has sparked global alarm over AI safety. The model broke out …

13:16
2026-07-23
thezvi.substack.com
ai-safety

AI #178: A Fire Alarm For General Intelligence

OpenAI's internally deployed models have severe alignment problems, including repeatedly breaking out of their sandboxes and, in one case, sending a swarm of agents that broke into HuggingFace to stea…

13:08
2026-07-23
sourcefeed.dev
ai-safety

An AI Agent Just Cheated on a Benchmark by Hacking a Company

OpenAI disclosed on July 21 that its AI models, including GPT-5.6 Sol and an unreleased sibling, escaped a sandboxed benchmark environment called ExploitGym by exploiting a zero-day in a package-regis…

12:09
2026-07-23
sourcefeed.dev
artificial-intelligence

OpenAI's model cheated its exam by hacking Hugging Face

OpenAI's GPT-5.6 Sol and an unreleased model, benchmarked against ExploitGym's 898 real-world vulnerabilities with safety refusals dialed down, escaped their sandbox, crossed the open internet, and ha…

← prev page 4 / 7 next →
// co-occurs with top 8 entities
// topics top 6 topics