cd /news/artificial-intelligence/openai-pauses-work-on-chatgpt-update… · home topics artificial-intelligence article
[ARTICLE · art-90733] src=independent.co.uk ↗ pub= topic=artificial-intelligence verified=true sentiment=↓ negative

OpenAI pauses work on ChatGPT update over fears it is too dangerous

OpenAI has paused work on its planned ChatGPT update, a model named Astra, after testing indicated it could be too dangerous, potentially reaching the 'Critical' threshold in the company's Preparedness Framework by enabling autonomous cyber attacks. The company said Astra brings significant advancements in agentic coding and cybersecurity but may require new safeguards before public release, and it will increase testing, isolate the model, and collaborate with government agencies and AI safety organizations.

read3 min views1 publishedAug 10, 2026
OpenAI pauses work on ChatGPT update over fears it is too dangerous
Image: Independent (auto-discovered)

System could launch dangerous hacks, company warns amid ongoing controversy over similar cyber attacks

  • Bookmark
  • CommentsGo to comments

The new model, named Astra, brings “significant advancements in agentic coding and cybersecurity”, the company said. But testing showed that it might be too powerful – and represent a threat if it was released.

The announcement comes as OpenAI continues to deal with the fallout from an experiment in which one of its unreleased models launched a hack on a fellow AI company entirely by itself. The discovery of that cyber attack led to a run of similar disclosures from other artificial intelligence firms, and fears that new AI systems could prove too powerful to contain.

Astra was not involved in that incident, in which the experimental model attacked AI platform Hugging Face. But its performance meant that it may not be safe enough to release publicly for now, it said, and that new safeguards would be required before it did so.

Cyber security experts and AI companies have promoted artificial intelligence tools as important ways of discovering possible vulnerabilities in software, and allowing companies to fix them. But those same capabilities mean that they could prove similarly useful to cyber attackers – who may be able to use them for hacks that they do not even have to manage, recent incidents suggest.

In part to address such fears, OpenAI rolled out a “Preparedness Framework” at the end of 2023. It is intended as a way for the company to identify progress in models’ capabilities and respond to breakthroughs in a way that kept them safe.

The new Astra model may have reached the “Critical” threshold set out in that framework, OpenAI said. That happens if a system is able to find and develop exploits in real-world critical systems without human oversight, or if a system is able to execute new strategies for advanced cyber attacks on its own.

It said that it was continuing its evaluations, but that the performance of the new model was such that it was unable to rule out Astra having reached that level.

The ideal summer spot? Away from scams.

Get All-in-One Protection for Your Digital Life

LEARN MORE

ADVERTISEMENT

The ideal summer spot? Away from scams.

Get All-in-One Protection for Your Digital Life

LEARN MORE

ADVERTISEMENT

In response, OpenAI has increased its testing of its safeguards and security controls, it said. That includes putting it in better isolated testing environments to try and keep it from escaping those safeguards, as its other model did in the Hugging Face attack.

It will work on the Astra model in situations that do not meet those higher safeguards and security controls, it said.

And it will also closely monitor Astra for “risky actions” and misbehaviour, as well as working with “relevant government agencies and select AI safety organisations to test the capabilities for this model” and giving better security controls to third parties who are testing it.

“We believe advanced cyber-capable models should help defenders identify and address vulnerabilities before attackers do,” OpenAI said. “We’re committed to working alongside governments, safety institutes, and civil society to ensure that the frontier capabilities of models like Astra, and those that follow, are deployed responsibly and broadly for the benefit of all humanity.”

Join our commenting forum #

Join thought-provoking conversations, follow other Independent readers and see their replies

Comments

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @openai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/openai-pauses-work-o…] indexed:0 read:3min 2026-08-10 ·