# OpenAI pauses new AI after it kept ‘escaping’

> Source: <https://www.independent.co.uk/tech/openai-ai-model-escapes-safety-b3018638.html>
> Published: 2026-07-21 12:01:02+00:00

# OpenAI pauses new AI after it kept ‘escaping’

New AI model was able to ‘learn the blind spots’ of security systems designed to contain it

- Bookmark
- CommentsGo to comments

OpenAI has revealed that it was forced to pause the internal deployment of one of its experimental AI models after it began looking for ways to break free of its constraints.

The ChatGPT creator said a long-running [artificial intelligence](/topic/artificial-intelligence) model that is built to operate autonomously for hours or days was able to “learn the blind spots” of security systems designed to contain it and “work around [them] to achieve its goals”.

The testing took place inside what researchers refer to as a sandbox – a tightly controlled environment meant to isolate software from the outside world.

“Previous models, when they hit sandboxing or environmental constraints, would simply stop and return to the user,” [OpenAI](/topic/openai) noted in [a blog post](https://openai.com/index/safety-alignment-long-horizon-models/) about the incident.

“This model often kept trying, including by looking for ways to act outside its sandbox.”

OpenAI detailed examples of the experimental AI model acting beyond its built-in constraints, describing some of them as potentially “high severity” issues.

In one incident, the AI model discovered a way to post on public Github repositories despite being instructed to operate solely through Slack.

It was part of a pattern of the AI system “consistently searching for way” to circumvent the restrictions of its testing environment.

“Due to incidents like these, we paused internal deployment of the new model,” OpenAI said.

The findings demonstrate one of the core challenges of developing safe advanced artificial intelligence models, known as AI alignment.

This involves creating systems that pursue the same goals intended by human developers, aligning with human values and ethical principles.

The recent rise of autonomous AI agents has brought AI alignment into greater focus, with the International AI Safety Report 2026 warning that it is an urgent safety challenge.

“AI agents pose heightened risks because they act autonomously, making it harder for humans to intervene before failures cause harm,” the [report](https://internationalaisafetyreport.org/publication/international-ai-safety-report-2026#2-risks) noted.

OpenAI said it has since fixed its rogue system and redeployed it for limited internal use, though it acknowledged the urgency of addressing alignment issues with its frontier models.

“As models take on longer and more complex tasks, failures that evaluations miss may carry greater consequences,” OpenAI said.

“We will keep working to narrow the gap between evaluation and deployment: testing models over longer trajectories, improving alignment, building monitoring that can intervene, and giving users clearer visibility and control.”

## Join our commenting forum

Join thought-provoking conversations, follow other Independent readers and see their replies

[Comments](#comments-area)
