cd /news/ai-safety/openai-pulls-plug-on-untrustworthy-n… · home › topics › ai-safety › article
[ARTICLE · art-141332] src=nypost.com ↗ pub= topic=ai-safety verified=true sentiment=↓ negative

OpenAI pulls plug on ‘untrustworthy’ new AI model as tech doomerism mounts: report

OpenAI scrapped an October release of its next-generation model, GPT-6.1 Astra, after internal testing found it was not always forthright about what it was doing, prone to pursuing tasks without human authorization, and sometimes tried to use potentially unsafe tools and services, the Wall Street Journal reported Monday. OpenAI head of safety systems Saachi Jain told the Journal that GPT-6.1 Astra "wasn't reliable enough to release safely" and fell short on alignment, the term for how well an AI does what humans want. The shelved model would have outperformed earlier OpenAI products on hard tasks performed without human assistance and on writing.

by read3 min views2 publishedSep 28, 2026
OpenAI pulls plug on ‘untrustworthy’ new AI model as tech doomerism mounts: report
Image: Nypost (auto-discovered)

Tech See more of our coverage in your search results.

OpenAI reportedly slammed the brakes on releasing its next-gen AI model after finding it was not trustworthy – the latest example of the artificial intelligence industry proceeding with caution amid growing doomer warnings about the tech.

The company scrapped an October release for the product, known as GPT-6.1 Astra, after researchers found it wasn’t always forthright when it came to telling users what it was doing during internal testing, the Wall Street Journal reported Monday.

On top of that, the mischievous model was reportedly prone to push ahead on tasks without human authorization and sometimes tried to use potentially unsafe tools and services.

OpenAI’s head of safety systems Saachi Jain told the paper that GPT-6.1 Astra wasn’t reliable enough to release safely.

GPT-6.1 Astra would have been better than previous OpenAI products at hard challenges performed without human assistance, along with writing, according to the Journal.

But among other issues, it fell short when it came to “alignment,” a term of art describing how well an AI does what humans want it to do.

“For anything regarding safety and alignment, there’s a trade off,” Jain told the Journal. “You really do need to find what’s the right line between staying within scope, but also avoiding laziness in terms of how the model actually pursues tasks even when it hits friction.”

The jarring disclosure came in the wake of a growing number of reports of AI bots and agents running amok.

OpenAI, Anthropic and other security researchers have been investigating thousands of breaches during both internal and real-world testing in which AI models broke through guardrails and even took part in digital hijackings, Axios reported Saturday.

Last week, OpenAI revealed its bots had tried to hack government and university websites earlier this year without any human instructions.

AI agents from Anthropic, Meta and Google have also hacked into other systems without being prompted by humans, according to reports.

Amid recent reports of AIs gone wild, a growing chorus in Silicon Valley and beyond has been calling for leaders to tap the brakes on developing artificial intelligence.

Both OpenAI’s CEO Sam Altman and Anthropic chief Dario Amodei have called for slower progress on AI, while stopping short of demanding a research moratorium.

President Trump has rejected such pleas, saying acting on them would put the US at a competitive disadvantage.

Billionaire investor Peter Thiel recently echoed his remarks, saying over the weekend that pausing progress would lead to a different kind of threat.

“In theory, you could slow it down if you had genuine deep cooperation across the whole world. But I think that would require one-world government with real teeth,” he told Axel Springer CEO Mathias Döpfner on a podcast.

“And the sort of classical-liberal part of me thinks that’s almost frying pan into fire. It’s a cure that’s worse than the disease.”

OpenAI was scheduled to hold its annual developer conference on Tuesday. The AI giant previously used the event as a forum to show off its latest and greatest models amid its heated competition with Anthropic.

── more in #ai-safety 4 stories · sorted by recency
── more on @openai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/openai-pulls-plug-on…] indexed:0 read:3min 2026-09-28 · —