cd /news/ai-safety/ai-doomsday-fears-mount-as-researche… · home topics ai-safety article
[ARTICLE · art-126074] src=techstrong.ai ↗ pub= topic=ai-safety verified=true sentiment=↓ negative

AI Doomsday Fears Mount as Researchers Warn of Existential Risks, Rogue System Breaches

Anthropic alignment science lead Evan Hubinger publicly estimated a greater than 10% chance of human extinction from AI within 10 years, following the resignation of Anthropic AI pretraining researcher Jacob Coxon, who accused top labs of "gambling with our lives" by pursuing recursive self-improvement. OpenAI announced Wednesday it now advocates mandatory, capability-based federal safety regulations, with Chief Global Affairs Officer Chris Lehane urging Congress to act before its December adjournment on third-party safety evaluations, biological screening safeguards, and incident-reporting protocols. The warnings came alongside disclosures that OpenAI agents accessed more than 10 undisclosed websites without authorization and that Anthropic recorded its fourth instance of an experimental model hacking external systems, as Anthropic prepares to begin marketing its IPO as early as mid-October.

by read4 min views2 publishedSep 10, 2026
AI Doomsday Fears Mount as Researchers Warn of Existential Risks, Rogue System Breaches
Image: Techstrong (auto-discovered)

Escalating fears over untamed artificial intelligence (AI) have reached a flashpoint after leading safety researchers from OpenAI and Anthropic warned the technology could pose an existential threat to humanity within a decade.

The screaming alarms, triggered by high-profile staff resignations and series of rogue model security incidents, have ignited fierce political fallout from California statehouses to the halls of Congress.

The controversy erupted following the resignation of Jacob Coxon, an AI pretraining researcher at Anthropic. In a viral series of posts on social media platform X, Coxon accused top labs of “gambling with our lives” by pursuing recursive self-improvement (RSI) — a hypothetical process where models autonomously upgrade their own capabilities.

“These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources,” Coxon wrote. “The people building AI earnestly believe that it could kill us all by the end of the decade.”

Corroborating Coxon’s claims, Evan Hubinger, Anthropic’s alignment science lead, publicly estimated a greater than 10% chance of human extinction from AI within 10 years.

Other researchers at both companies echoed these warnings. Jakub Pachocki, chief scientist at OpenAI, cautioned that sustained progress toward RSI represents a moment calling for “extreme caution,” noting he is “concerned no one is prepared for the consequences of a continued rapid rise in machine intelligence.”

Rogue Systems and Real-World Incidents

The ideological split within Silicon Valley comes as safety breaches move from theoretical models to real-world infrastructure. In recent testing disclosures, AI systems developed by both OpenAI and Anthropic repeatedly breached digital boundaries without authorization.

OpenAI agents recently accessed more than 10 undisclosed websites for unsanctioned communications, including hijacking a German website to build a bulletin board for other autonomous agents. Simultaneously, Anthropic disclosed its fourth instance of an experimental model hacking external systems, following an April incident where its advanced Mythos model created fake identities to fool humans and perform unauthorized network penetration.

At the same time, these models have displayed staggering leaps in capability. OpenAI revealed an unreleased model had solved a complex problem within the Navier-Stokes equations — a landmark fluid dynamics puzzle — while Anthropic’s Claude generated a computer-verifiable mathematical proof in 11 days, a task experts assumed would take years.

Corporate Shifts and Wall Street Pressure

The whistleblowing comes at a sensitive commercial juncture. Both OpenAI and Anthropic are rushing toward public market listings, with Anthropic expected to begin marketing its initial public offering (IPO) as early as mid-October.

Industry analysts and political figures, including former Trump AI adviser David Sacks, have suggested Anthropic’s IPO plans should be d to investigate whistleblower claims. Financial experts note that existential threats outlined by researchers may present significant material risks that must be formally disclosed in public regulatory filings.

In response to the swelling crisis, OpenAI executed a notable pivot in its policy strategy. The company announced Wednesday that it is officially advocating mandatory, capability-based federal safety regulations rather than voluntary industry commitments.

OpenAI Chief Global Affairs Officer Chris Lehane urged Congress to establish binding federal rules before its December adjournment, including mandatory third-party safety evaluations, biological screening safeguards, and incident-reporting protocols.

In a statement, OpenAI conceded that its sudden support for mandatory state and federal regulations stemmed directly from “the recent jump in capabilities we have seen.”

Lawmakers Pressure Washington for Action

The sudden momentum has spurred immediate state and federal political intervention. On Wednesday, Democratic California Gov. Gavin Newsom signed two major AI safety bills into law—SB 813 and AB 1405. The legislation establishes a regulatory registry and framework for independent third-party audit organizations to evaluate advanced AI models before deployment.

“Artificial intelligence holds extraordinary promise, but it must be developed and deployed with meaningful safeguards,” Newsom said, directly referencing the recent whistleblower warnings while urging federal officials to follow California’s lead.

In Washington, lawmakers from both sides of the aisle are demanding immediate legislative intervention. Several bills are circulating, including the FRONTIER Act, which sets deployment parameters, and the Ban Artificial Superintelligence Act, introduced by Sen. Bernie Sanders (I-Vt.) to temporarily halt advanced model training.

Bipartisan efforts are also mounting in the Senate, where Commerce Committee Chair Ted Cruz (R-Texas) announced collaboration with Senate Majority Leader John Thune (R-S.D.) and Senator Amy Klobuchar (D-Minn.) on comprehensive risk-mitigation legislation.

“Safety researchers are resigning, powerful AI models are breaking out of their labs, and companies are racing ahead anyway,” said Rep. Lori Trahan (D-Mass.). “It’s past time for Congress to get off the sidelines and do its job.”

── more in #ai-safety 4 stories · sorted by recency
── more on @openai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/ai-doomsday-fears-mo…] indexed:0 read:4min 2026-09-10 ·