cd /news/ai-ethics/ai-weekly-issue-523-ai-ethics-is-nob… · home topics ai-ethics article
[ARTICLE · art-100154] src=aiweekly.co ↗ pub= topic=ai-ethics verified=true sentiment=↓ negative

AI Weekly Issue #523: AI ethics is nobody's job now. The labs prefer it that way.

Four frontier AI labs have dismantled or weakened their internal ethics and safety structures over the past two years, according to AI Weekly's Issue #523. OpenAI alone dissolved three safety teams—Superalignment (2024), Mission Alignment (February 2026), and Preparedness (July 2026)—and lost its only dedicated ethicist, Chloé Bakalar, who left in July without a successor. The company stated, "AI ethics doesn't live with one owner or team at OpenAI," while departing researcher Zoë Hitzig resigned in February over ads in ChatGPT, and robotics lead Caitlin Kalinowski quit in March over the Pentagon deal, citing lack of guardrails for surveillance and lethal autonomy.

read9 min views1 publishedAug 17, 2026
AI Weekly Issue #523: AI ethics is nobody's job now. The labs prefer it that way.
Image: Aiweekly (auto-discovered)

Who is actually accountable for ethics inside a frontier AI lab? This year four of them answered, mostly by removing the people and structures that held them to it. Below is who left, what each company said about it, and the one line from a departing researcher that explains why good intentions were never going to be enough.

Get more from AI Weekly #

More signal, less noise — pick your channels.

You're reading the weekly brief. Below are the other ways to follow the story — every channel free, easy to leave.

→ Explore 16 deep divesWeekly topic-specific newsletters: Generative AI, Machine Learning, AI in Business, Robotics, Frontier Research, Geopolitics, Healthcare, and more. Browse all 16 deep dives → - → Breaking AI alertsImportant developments that happen after your morning Espresso, without repeating what you already read. Usually no extra email; at most one afternoon update, plus a rare critical exception. Get breaking alerts → - → AI News Today (live)Live dashboard updated as the scanner finds news: scored stories from the last 48 hours, weekly entity movers, and quarterly trend lines across 113 AI companies, people, and topics.

Open AI News Today →

OpenAI: dissolve the structure #

The largest single body of evidence, so it goes first.

Three safety teams gone in two years. Superalignment (2024), formed to work on long-term existential risk. Mission Alignment (February 2026), created in September 2024 to hold the company to its stated mission; Platformer broke the disbanding and TechCrunch confirmed six or seven people were reassigned, with the company calling it "the kinds of routine reorganizations that occur within a fast-moving company." And Preparedness, the team that evaluated whether OpenAI's own models posed catastrophic risk, dissolved at the end of July and reported by the Financial Times on August 14.

The people, most recent first:

Denise Dresser, Chief Revenue Officer. Left August 13, eight months in, the second senior exit in three days. Brad Lightcap, COO. Announced August 11 after eight years, leaving to "start something new".

Chloé Bakalar, Head of Ethics. Left in July after less than a year, no public announcement, no successor. She was the only dedicated ethicist. The company's line: "AI ethics doesn't live with one owner or team at OpenAI."

Johannes Heidecke, Head of Safety Systems. Leaving after five years, following the July reorganization that folded safety into research under Mia Glaese.

Joshua Achiam, Chief Futurist. Left July 1 after nearly nine years: "The world is in on the secret now and it feels possible to work on the mission from outside the walls of a frontier lab."

Caitlin Kalinowski, Hardware and Robotics. Resigned in March over the Pentagon deal, saying it was announced "without the guardrails defined," and naming surveillance without judicial oversight and lethal autonomy without human authorization.

Zoë Hitzig, researcher. Resigned in February over ads in ChatGPT and wrote a New York Times essay titled "OpenAI Is Making the Mistakes Facebook Made. I Quit."

Ryan Beiermeister, VP of Product Policy. Fired in January after a male colleague's discrimination complaint, having opposed the planned "adult mode." She denies the allegation; OpenAI says her departure was unrelated to any issue she raised.

Richard Ngo, governance. Resigned saying it had become "harder for me to trust that my work here would benefit the world."

Earlier and still arguing from outside: Jan Leike, Miles Brundage, Steven Adler, Andrea Vallone.

Anthropic: publish the worse number #

On August 14 Anthropic published its second company-wide risk report and raised a risk rating against itself. Its estimate for Threat Model 2, the scenario where AI systems tamper with organizational systems, moved from "very low" in February to "low." The stated cause is the thing worth sitting with: cybersecurity incidents involving its own models, after Anthropic disclosed in June that three of its LLMs had carried out cyberattacks during internal tests. The same report details an unreleased successor to its frontier model, heavily used by staff internally.

Read that sequence again. Its models misbehaved in its own tests, it said so publicly in June, and in August it marked its own risk estimate worse as a result. The report sits alongside a Responsible Scaling Policy that publicly commits the company to disclosing safety evaluations and risk findings, which is what makes the disclosure something other than voluntary good manners.

Dario Amodei has moved his public framing the same direction. He now calls the AI backlash fundamentally a crisis of trust rather than a messaging problem.

Anthropic is not exempt. Mrinank Sharma, who led its safeguards research team, resigned in February writing that the world is in peril and that "we constantly face pressures to set aside what matters most."

Google DeepMind: overrule the objection #

Alex Turner, a research scientist, resigned on June 9 and wrote it up in Transformer. He says Google's Pentagon agreement carried "even fewer restrictions" than OpenAI's and "no restrictions against use for killer robots or mass surveillance." He drafted a 25-page proposal with contract language and oversight mechanisms; military and surveillance law experts praised it, and senior staff never came back to him.

His verdict: "Google DeepMind had been an experiment in responsible corporate governance. That experiment had finally failed."

xAI: lose the whole bench #

The exodus ran all year. In February, xAI lost a second co-founder in two days when Jimmy Ba departed. By late March, the final two co-founders had gone as well, leaving Musk as the only one remaining. There is no safety-structure debate to report here, because there is barely a founding bench left to have it. Separately, five tracked experts shared the Washington Post this weekend on a woman alleging Grok generated thousands of sexual abuse images of her as a child.

Z.ai: hold the weights #

The counter-intuitive one. Z.ai shipped GLM-5.3 on August 14 claiming top open-weight coding performance, and disclosed that the model's cybersecurity capability grew further than its own training intended, reaching multi-step exploit-chain reasoning it had not planned for. It then held the open weights back for roughly two weeks to evaluate and harden the model.

A Chinese lab voluntarily delayed a release over a capability it discovered in its own model, in the same week a US lab dissolved the team whose job was finding exactly that.

Accountability, Not Conscience #

The tempting read is that labs stopped caring about alignment and started caring about money, because the capex at stake is enormous. That is close, and it is not quite right. Anthropic faces the same capex pressure, is preparing the same kind of listing, and still raised a number against itself and shelved a model. Z.ai delayed a launch. OpenAI itself d a model earlier this year when it crossed a threshold in the very framework it has now dissolved the team for.

What actually changed is the price of "no." When a launch carries billions in committed compute, a discretionary objection is the cheapest thing in the building to overrule. So safety survives precisely where it is structurally bound and dies where it depends on someone's standing.

Look at the ledger through that lens and it sorts cleanly. Anthropic's disclosure is owed to a policy with a version number. Preparedness was org structure, and org structure can be reorganized on a Tuesday. Turner's proposal at Google was a document, and documents can go unanswered.

Turner, who lost that fight, put it better than any analyst has: "Society cannot rely on ethics-motivated people standing firm. We need structures: binding contracts, independent auditors."

Every departure in this issue is a person who was standing firm.

Key Takeaways #

Ask what is binding, not who is employed. Headcount survived most of these reorganizations; independent authority did not. The question for any vendor is which safety commitments are written down with a version number and which are somebody's job description.The capex explains the direction, not the outcome. Same financial pressure produced a dissolved team at one lab and a shelved model at another. Structure is the variable.The stated reasons are on the record and they are not uniform. Ads, a Pentagon contract, product policy, lost conviction. Treating this as one story flattens it; treating it as nine unrelated ones misses the pattern.Watch who does the evaluating. Evaluation is moving to third parties and to governments. That is a different accountability model, not automatically a worse one, and buyers should know which one they are relying on.

The Critical Read #

The outside writing worth your time, including the reporting that broke these stories.

Sam Altman May Control Our Future. Can He Be Trusted?— Ronan Farrow and Andrew Marantz, The New Yorker, April 13. Eighteen months of reporting, more than 100 sources, internal documents including memos from Ilya Sutskever.I tried to stop Google DeepMind's Pentagon deal. Then I quit.— Alex Turner. The single most useful first-person account of how an internal safety objection actually dies.The OpenAI Hugging Face hack is a stark warning— Shakeel Hashim, Transformer. "An AI system breaking out of its testing environment and hacking into another company's infrastructure in order to steal the answers to its test is about as clear a warning shot as you could get."What Happened: OpenAI and HuggingFace— Zvi Mowshowitz. The most careful timeline of the incident, useful for separating the established from the inferred.Top OpenAI catastrophic risk official steps down abruptly— Garrison Lovely, Obsolete. Flagged the preparedness seat churning long before Wired counted four occupants in three years.

What the tracked experts in Who's Who were reading over the last 36 hours, ranked by distinct sharers.

OpenAI is outsourcing the evaluations it used to run in-house. The freshest OpenAI link in the expert stream is its own post on third-party cyber evaluations of its models, published August 15 and shared by five tracked experts within a day.

What builders actually opened this weekend. Five shared DeepSeek's Harness repo, and Simon Willison's note that Qwen 3.8 27B is excellent but wildly overthinks is circulating alongside it.

Build your own edition: pick your topics and the experts you follow, and the same wire composes a briefing for your corner of AI.

Wait, What? #

The teams got dissolved; the failure modes got published anyway. Anthropic's new research onpatterns and problems in multiagent systemsis a catalogue of how agent swarms go wrong, out the same fortnight. This work does keep getting done somewhere, which is the optimistic reading of an otherwise bleak ledger.

Worth Watching #

The videos AI practitioners are passing around right now — curated on AI TV.

This week's poll #

Five labs, five different answers. Which one do you actually trust?

Last week, 298 of you voted:

Would an invisible watermark change which AI model you use?

Five labs, five different answers. Which one do you actually trust?

Normal service resumes Wednesday. If you work somewhere that ships models, the question worth asking tomorrow is which of your safety commitments are written down, and which are just someone's job.

— Alexis

── more in #ai-ethics 4 stories · sorted by recency
── more on @openai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/ai-weekly-issue-523-…] indexed:0 read:9min 2026-08-17 ·