Midterms 2026See who we think should earn your vote, based on our standardsThe guide →
WRITTEN IN PLAIN AMERICAN ENGLISH.
CLAY TRIBUNE.
Advertisement

Anthropic’s CEO Pushes Outside Oversight — From People He Knows Well

Anthropic CEO Dario Amodei wants independent 'embedded evaluators' to watch AI companies, including his own. A look at the safety community behind the plan.

By mitch·5 min read
A glowing AI robot head sits on a desk surrounded by books, symbolizing oversight and evaluation.

Anthropic’s CEO wants outside watchdogs to watch the company’s peers, and the watchdog he cites comes from the same safety community that birthed his own company.

Dario Amodei, CEO of Anthropic, has proposed that independent “embedded evaluators” supervise AI development companies. These evaluators would verify safety practices, report incidents, and assess model alignment and training pipelines. Amodei wrote: “We are still the ones choosing what to include and omit. Embedded evaluators will change this dynamic.”

The proposal puts the safety community at the center of its own oversight. That is the central tension in this story: a group that built its reputation on caution is now asking its own members to judge it.

Advertisement

The Man Behind the Plan

Amodei is a physicist turned AI executive. He earned a Ph.D. in biophysics from Princeton in 2011 and spent years building neural networks at major tech companies before founding Anthropic.

His resume reads like a tour of Silicon Valley’s AI pipeline. At Baidu, he worked on speech recognition through machine learning. At Google, he helped develop neural-net research. At OpenAI, he became vice president of research during the development of ChatGPT 2 and ChatGPT 3.

He left OpenAI in 2020 because he believed the company wasn’t installing guardrails on the technology. He also didn’t know if he could trust the company to set aside financial interests.

“When you feel that you can’t trust someone, when you feel that their values are not what they say they are, when you feel that they’re not honest, when you feel that they’re not in it for the reasons that they say, when you see disturbing patterns of behavior, dishonesty, that makes it very hard to continue to wo”

That mistrust was part of why he left.

The Watchdog He Cites

The organization Amodei points to is METR, an AI safety testing laboratory. Its founder and CEO is Beth Barnes, who worked at OpenAI alongside Amodei during early ChatGPT development.

METR’s other co-founder is Paul Christiano. He led research around OpenAI’s model strategy controls and founded the first iteration of METR. He has framed his approach as effective altruism.

Barnes and Christiano have also provided safety checks for Anthropic’s AI Claude.

The Funding Circle

Amodei’s path to Anthropic passed through a funding circle that now defines the AI governance debate.

Jaan Tallinn led Anthropic’s 2021 Series A financing round. He is a prominent supporter of the Effective Altruism movement. Sam Bankman-Fried led Anthropic’s 2022 Series B financing round before FTX collapsed and he was convicted in a multibillion-dollar fraud case.

The Proposal in Detail

Amodei’s plan calls for evaluators embedded within companies to monitor AI development. Their role would be verification, reporting, and assessment.

The proposal is explicit about what it changes. Amodei wrote that the evaluators would alter “the dynamic” of who chooses what gets included and omitted.

The Safety Community

The safety community grew from the same people who built OpenAI’s early models. Christiano’s work on model strategy controls at OpenAI laid the groundwork for the approach METR now follows.

Barnes’s role at OpenAI placed her directly in the room where ChatGPT was developed. Her later work at METR connects the two organizations.

The overlap is notable. The person who tested Anthropic’s Claude also tested OpenAI’s early GPTs.

Why the Proposal Matters

Amodei’s proposal is a direct attempt to address the problem he saw at OpenAI. He left that company because he believed it wasn’t installing guardrails on the technology, and he now wants the industry to build those guardrails elsewhere.

Whether an evaluator from the same community can truly be independent is a fair question. The source does not answer it.

What We Know

  • Dario Amodei is CEO of Anthropic.
  • He proposes independent “embedded evaluators” to supervise AI development companies.
  • METR is the organization he cites as an example of the guidance he believes the industry needs.
  • Beth Barnes is METR’s founder and CEO; she worked at OpenAI alongside Amodei during early ChatGPT development.
  • Paul Christiano led research around OpenAI’s model strategy controls and founded the first iteration of METR; he has framed his approach as effective altruism.
  • Sam Bankman-Fried led Anthropic’s 2022 Series B financing round before FTX collapsed and he was convicted in a multibillion-dollar fraud case.
  • Jaan Tallinn led Anthropic’s 2021 Series A financing round; he is a prominent supporter of the Effective Altruism movement.
  • Barnes and Christiano provided safety checks for Anthropic’s AI Claude.
  • Amodei originally studied biophysics, earning a Ph.D. from Princeton in 2011.
  • At Baidu, he worked on speech recognition through machine learning.
  • At Google, he helped develop neural-net research.
  • At OpenAI, he became vice president of research during the development of ChatGPT 2 and ChatGPT 3.
  • Amodei left OpenAI in 2020 because he believed the company wasn’t installing guardrails on the technology.
  • He also doubted whether he could trust the company to set aside financial interests.

Where the paper stands

The paper backs narrow rules forcing companies to disclose safety failures they hid and is against broad rules that hand the market to the incumbents, and it is against letting big companies hide their failures while asking for oversight over others. In this case, the proposal’s central flaw is that the evaluators come from the same community whose judgment is at stake. An evaluator who has tested both Anthropic’s Claude and OpenAI’s early GPTs is hardly independent.

This is a familiar pattern. Big tech asks for oversight of its peers while keeping its own failures hidden. The paper has said before that licensing regimes and compliance costs only giants can afford are a moat, not a safeguard. Here the moat is the same safety community itself, and the proposal does nothing to breach it.

The reader should watch for the moment when the evaluators stop being observers and start being defenders of the status quo. That is when the dynamic changes for real, and it is not the change Amodei promises.

Source material: “Who is Dario Amodei, the Anthropic CEO pushing independent AI oversight?,” Fox News.

The Notebook

Get the Notebook.

The day's best stories and every fresh verdict, in plain English, in your inbox by seven. One email a day, no more.

We send one note to confirm. Every issue has a one-click way out.

Advertisement

Leave a Reply

Your email address will not be published. Required fields are marked *

As an Amazon Associate, Clay Tribune earns from qualifying purchases.