cd /news/ai-safety/what-on-earth-is-an-embedded-evaluat… · home topics ai-safety article
[ARTICLE · art-128963] src=businessinsider.com ↗ pub= topic=ai-safety verified=true sentiment=↑ positive

What on earth is an 'embedded evaluator'? One of the most important jobs in AI, according to frontier lab chiefs.

Anthropic CEO Dario Amodei called on frontier AI labs to embed independent safety evaluators inside their organizations in a blog post on Saturday, granting them employee-like access, badges, laptops, and the right to publish findings without company editorial control. OpenAI CEO Sam Altman endorsed the idea and said OpenAI "will do the same," while SpaceXAI CEO Elon Musk reposted Amodei's post with "Dario is right." AI Verification and Evaluation Research Institute executive director Miles Brundage said embedded auditors are a "critical part of the package" but insufficient alone, arguing the industry needs binding requirements so auditors are not selected and paid by the companies they audit.

by read3 min views1 publishedSep 14, 2026
What on earth is an 'embedded evaluator'? One of the most important jobs in AI, according to frontier lab chiefs.
Image: Businessinsider (auto-discovered)

The next high-profile AI hire may not be another researcher, but an

"embedded evaluator" tasked with scrutinizing frontier models before they're released. In a blog post on Saturday, Anthropic CEO Dario Amodei said frontier AI labs should commit to embedding independent safety evaluators within their organizations.

The embedded evaluators' job is to check whether the company "is actually following the training, deployment, operational, and safeguards practices they claim to be following," Amodei wrote.

He said embedded evaluators will have "employee-like access to verify safety practices and report incidents." They will have desks in the Anthropic offices, access badges, and company laptops, as well as the right to publish any findings without Anthropic's editorial control.

Amodei's plan comes as fears of an AI apocalypse reach a fever pitch, and it received an outpouring of support, even from executives he's feuded with.

OpenAI CEO Sam Altman, reposting Amodei's X post, wrote: "Committing to having independent evaluators with employee-like access is a great idea, and we will do the same."

SpaceXAI CEO Elon Musk also reposted Amodei's post, adding: "Dario is right." The idea has also received some VC attention. Sriram Krishnan, a former Andreessen Horowitz partner and former AI advisor to President Donald Trump, spoke about the importance of a distributed network of evaluators.

"The more eyes and people with distributed skill sets the better," Krishnan said in a Saturday X post. "It would be a good idea to fund several efforts on this."

Top AI talent is migrating to this space #

Amodei already has candidates in mind for the new job. In his post, he mentioned Berkeley-based Metr, a prominent nonprofit AI watchdog that conducts independent evaluations of AI models.

Metr, established in 2022 by ex-OpenAI staffer Beth Barnes, is attracting top talent from the biggest AI labs.

Joe Benton, previously a member of Anthropic's safety and oversight team, announced on Friday that he had left the company to join Metr. Josh Engels, a former employee of Google DeepMind's AGI safety team, said on Sunday that he had resigned and joined Metr because of the high stakes of AI safety.

Meanwhile, research labs are offering themselves up for the role of embedded evaluations. Christopher Manning, a senior fellow at Stanford's Institute for Human-Centered AI and the founder of Stanford's Natural Language Processing Group, said the group would be best suited for the job.

"For important parts of the work, universities would be better than any other organization," Manning wrote in an X post on Saturday.

Embedded evaluators aren't the golden ticket out of an AI apocalypse #

AI safety experts agree that embedded evaluators are important, but they also have limitations.

Miles Brundage, the executive director of the San Francisco-based think tank, the AI Verification and Evaluation Research Institute, told Business Insider that embedded auditors aren't sufficient on their own, but they're a "critical part of the package." Brundage was formerly an OpenAI senior advisor.

Brundage said the industry needs "binding requirements" to prevent auditors from being beholden to their host companies, and they should ideally not be selected and paid by the companies they audit.

"But companies can and should get started today," Brundage added.

Embedded evaluators will be most effective if they have a way to report potentially illegal behavior to an external safety committee unaffiliated with the AI labs they work in, said Kevin Frazier, a professor at the University of Texas School of Law who leads its AI Innovation and Law program.

Frazier proposed that evaluators should be embedded in AI labs for staggered, overlapping 26-month terms, "roughly the deployment of two new model classes," which would mean that they can assess how a lab has corrected prior errors in a new release.

He said the short window also prevents them from getting too connected with the lab's employees or culture.

"To be blunt, this will help make sure they do not drink the Kool-Aid," Frazier added.

── more in #ai-safety 4 stories · sorted by recency
── more on @anthropic 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/what-on-earth-is-an-…] indexed:0 read:3min 2026-09-14 ·