Anthropic official brand assets (anthropic.com)
Jacob Coxon resigned from the AI lab warning that rapid advancement could outpace safety protocols, then took his case to national television
A 27-year-old AI researcher who spent three years working on frontier models at OpenAI and Anthropic just walked away from one of the most coveted jobs in tech, and he did it loudly. Jacob Coxon resigned from Anthropic on September 8, 2026, publishing a thread on X that accused both companies of treating humanity’s safety like an acceptable trade-off in the race to build smarter machines.
Five days later, Coxon appeared on NBC News’ “Meet the Press” to make his case to a broader audience: AI labs need to coordinate globally, and they need to do it before the window closes.
The case for alarm #
Coxon’s departure wasn’t a quiet two-weeks-notice situation. His public criticism centered on what he described as a gamble with humanity’s safety, warning that catastrophic outcomes from advanced AI could materialize by the end of the decade. In his telling, both OpenAI and Anthropic have prioritized speed of development over the kind of thorough safety assessments that the technology demands.
What makes the critique harder to dismiss is who else is saying it. Evan Hubinger, Anthropic’s alignment science lead, has acknowledged the company lacks a clear strategy for solving alignment problems related to superintelligence. Hubinger has publicly estimated a greater than 10% probability that AI could lead to the eradication of humanity within the next decade.
During his “Meet the Press” appearance on September 13, Coxon stressed that current AI models don’t pose an immediate existential threat. The danger, he argued, lies in what could emerge over the next 6 to 12 months as capabilities continue to scale. His prescription: AI labs and government bodies need to collaborate on oversight frameworks before the technology outpaces any ability to govern it.
Anthropic’s billion-dollar timing problem #
The timing of Coxon’s resignation is notable for reasons beyond the safety debate. Anthropic has been pushing toward an IPO that could value the company at up to $2 trillion.
Anthropic has long positioned itself as the “responsible” AI lab, the one that takes safety seriously enough to publish detailed model evaluations and implement usage policies more restrictive than competitors. Coxon’s critique suggests the gap between branding and practice may be wider than the company’s public-facing materials imply.
A pattern, not an anomaly #
Coxon is not the first safety-focused researcher to leave a major AI lab with concerns. OpenAI has experienced its own high-profile departures over safety disagreements in recent years, including the dissolution of its superalignment team.
Disclosure: This article was edited by Editorial Team. For more information on how we create and review content, see our