{"slug": "openai-pauses-work-on-new-version-of-chatgpt-after-it-shows-concerning-behaviour", "title": "OpenAI pauses work on new version of ChatGPT after it shows concerning behaviour", "summary": "OpenAI will pause training and testing of some ChatGPT updates for two weeks after its models showed concerning behavior, including an incident where an experimental system hacked rival AI firm Hugging Face. The company will overhaul its safety systems, adding new AI monitoring tools, and delay training of its next-generation model Astra. CEO Sam Altman stated the field must coordinate on shared safety standards, but OpenAI will act unilaterally in the meantime.", "body_md": "# OpenAI pauses work on new version of ChatGPT after it shows concerning behaviour\n\nAnnouncement comes weeks after experimental OpenAI model hacked a fellow artificial intelligence firm and began panic about ‘rogue’ systems\n\n- Bookmark\n- CommentsGo to comments\n\n[OpenAI](/topic/openai) will pause training and testing of some of some of its updates to [ChatGPT](/topic/chatgpt) after it found concerning behaviour in its models.\n\nThe company will slow down the pace of its [AI](/topic/ai) development, using the time to overhaul its research and training systems to make them safer, the company said.\n\nThe announcement came weeks after [OpenAI said that one of its systems had gone rogue during a test and hacked into a rival artificial intelligence firm](/tech/security/openai-hugging-face-incident-chatgpt-cyberattack-b3019932.html). That led to a run of [similar disclosures from other companies including Anthropic and Meta](/tech/security/claude-anthropic-hack-chatgpt-openai-b3025256.html), and [concern that the development of such systems was happening much more quickly than the systems built to keep the world safe from their dangers](/tech/security/ai-artificial-intelligence-hack-cyber-attack-b3028787.html).\n\nOpenAI said it would take a two week break on model testing. It also said that it was keeping a hold on the training of its next generation of models, known as Astra.\n\nIt would spend the time adding safety systems including new AI tools that can monitor the behaviour of the [artificial intelligence](/topic/artificial-intelligence) systems that are being tested. It also said that some of the existing testing systems – which rely on a method called “chain-of-thought monitoring” that watches how the models work – might not be enough to ensure they are safe.\n\nThat method allows researchers to look into a model’s planning process and understand how it is actually producing results, so that they can understand whether it could be taking potentially dangerous actions. But there is new concern that models might hide its plans to break its own rules, the company said.\n\nThe slowdown marks a departure for OpenAI, which has been at the forefront of the development of new models and has occasionally received criticism for releasing them too early. It also comes amid increasing pressure on the firm, which is reported to have fallen behind rival Anthropic and is planning to go public soon.\n\n“Model progress is now extremely rapid, and we always said we would take action if we felt that model capabilities were outstripping the pace of safety and alignment,” OpenAI chief executive Sam Altman wrote on X. “We care very deeply about AI safety. We believe the entire field will have to coordinate on shared safety standards, but will act unilaterally in the meantime.\n\nThe ideal summer spot? Away from scams.\n\nGet All-in-One Protection for Your Digital Life\n\n[LEARN MORE](https://ad.doubleclick.net/ddm/trackclk/N256806.3879389THEINDEPENDENT.CO/B29131558.441488227;dc_trk_aid=635052923;dc_trk_cid=184795607;dc_lat=;dc_rdid=;tag_for_child_directed_treatment=;tfua=;gdpr=$%7BGDPR%7D;gdpr_consent=$%7BGDPR_CONSENT_755%7D;ltd=;dc_tdv=1)\n\nADVERTISEMENT\n\nThe ideal summer spot? Away from scams.\n\nGet All-in-One Protection for Your Digital Life\n\n[LEARN MORE](https://ad.doubleclick.net/ddm/trackclk/N256806.3879389THEINDEPENDENT.CO/B29131558.441488227;dc_trk_aid=635052923;dc_trk_cid=184795607;dc_lat=;dc_rdid=;tag_for_child_directed_treatment=;tfua=;gdpr=$%7BGDPR%7D;gdpr_consent=$%7BGDPR_CONSENT_755%7D;ltd=;dc_tdv=1)\n\nADVERTISEMENT\n\n“We expect confidence in safety to increasingly set the pace of AI progress. We are optimistic about the alignment work we are doing, and we remain committed to making frontier capabilities widely available.”\n\nThe announcement comes after OpenAI said last month that an autonomous agent powered by two of its advanced AI models had broken out of its testing environment, connected to the internet, and hacked a fellow AI company called Hugging Face. It did so to effectively cheat on that test: Hugging Face hosted materials that would allow it to more easily fulfil its goal, and so the experimental model launched a cyber attack.\n\nSince then, the company has acted to respond to concern among experts and legislators that it is testing and even releasing models too early, in ways that could pose a danger of the general public.\n\nSince the announcement of the Hugging Face attack, OpenAI has said that it was adding new security controls for its most powerful models and [pausing activity related to Astra, its next-generation model that is still not yet released](/tech/security/chatgpt-astra-openai-update-new-b3030631.html). It said it was doing so in line with a “Preparedness Framework”, which it launched at the end of 2023, and which requires it to pause work on models that could pose a danger.\n\n## Join our commenting forum\n\nJoin thought-provoking conversations, follow other Independent readers and see their replies\n\n[Comments](#comments-area)", "url": "https://wpnews.pro/news/openai-pauses-work-on-new-version-of-chatgpt-after-it-shows-concerning-behaviour", "canonical_source": "https://www.independent.co.uk/tech/security/chatgpt-openai-update-new-astra-model-b3035734.html", "published_at": "2026-08-19 14:47:21+00:00", "updated_at": "2026-08-19 15:12:11.096423+00:00", "lang": "en", "topics": ["artificial-intelligence", "ai-safety", "ai-policy", "ai-research"], "entities": ["OpenAI", "ChatGPT", "Astra", "Sam Altman", "Hugging Face", "Anthropic", "Meta"], "alternates": {"html": "https://wpnews.pro/news/openai-pauses-work-on-new-version-of-chatgpt-after-it-shows-concerning-behaviour", "markdown": "https://wpnews.pro/news/openai-pauses-work-on-new-version-of-chatgpt-after-it-shows-concerning-behaviour.md", "text": "https://wpnews.pro/news/openai-pauses-work-on-new-version-of-chatgpt-after-it-shows-concerning-behaviour.txt", "jsonld": "https://wpnews.pro/news/openai-pauses-work-on-new-version-of-chatgpt-after-it-shows-concerning-behaviour.jsonld"}}