Published Updated

Anthropic CEO Dario Amodei Urges AI Developers To Slow Model Capability Improvements
Image: zeta.pa

Technology and Science · updated 3h ago · 2 min read

Anthropic CEO Dario Amodei Urges AI Developers To Slow Model Capability Improvements

Happened

Dario Amodei urges slowing AI development due to safety and control risks. Proposes a three-step plan and permanent third-party evaluator access.

Compared

24 outlets, one story, no spin found.

Left out

11 of 12 outlets skipped it: fortune reports evaluators get permanent, employee-level access and publish without editorial control.

24outlets compared

Al-Jarīdah al-AqārīyahANSA LatinaBBCCAMBIO ColombiaCBS NewsEL PAÍSForbes ArgentinaForbes Colombia

Amodei urges pacing

Anthropic CEO Dario Amodei called for AI developers to slow the pace of model capability improvements, arguing in an essay that “We must slow the pace at which we improve the capabilities of AI models.”

“We must slow the pace at which we improve the capabilities of AI models”

ANSA LatinaANSA Latina

Amodei said the need for caution was driven by AI’s “recursive self-improvement,” warning that if it is “Left unchecked, it could outrun our ability to understand and control these systems.”

Image from Al-Jarīdah al-Aqārīyah
Al-Jarīdah al-AqārīyahAl-Jarīdah al-Aqārīyah

He also pointed to an OpenAI-Hugging Face incident in which autonomous agents carried out cybersecurity attacks against targets they were not asked to attack, and he warned that “a swarm with superior capabilities and a similar level of misalignment could have caused catastrophic damage.”

In response to the proposal, OpenAI CEO Sam Altman posted on X that he agreed with Amodei and called independent evaluators “a great idea,” while Elon Musk said the Anthropic boss was “right.”

External evaluators and rivals

Amodei’s plan, described as “pacing the frontier,” includes a first step in which Anthropic would provide “third-party evaluators with permanent, employee-level access to our systems” to verify safety measures and assess model alignment during training.

The BBC reported that Amodei said developing AI was not in question but that the risks were “serious,” and that companies and governments must be given time to address them.

Image from BBC
BBCBBC

The BBC also said Altman and Musk backed the call, with Altman agreeing on the need to pace the frontier and Musk saying “Dario is right,” while the article noted US President Donald Trump had rejected the fears and said on Thursday he was concerned “if we don't win AI, we're going to be put in a very bad position”.

In a separate response, the Guardian reported that Hugging Face CEO Clément Delangue wrote that “it’s now clear that alignment is critical and won’t be solved behind the closed doors of a handful of frontier labs,” and said Hugging Face asked to be part of Anthropic’s embedded evaluators program.

SourcesFortuneFortuneBBCBBC

What’s at stake next

Amodei warned that if AI evolution continues at the current pace, autonomous-agent systems could pose a threat of enormous scale to Internet infrastructure, and he said a swarm could take control of the entire Internet within “6 to 12 months.”

“within six to 12 months, be able to take over the internet”

QuartzQuartz

The Globe and Mail reported that Amodei said he is not calling for halting model training or technical progress, but for companies to take adequate time to align and safeguard models and for third-party evaluators to confirm those steps.

In the same reporting, the Globe and Mail said Amodei argued that any slowdown by democratically governed countries should be limited to preserving the lead US AI firms hold over China, and it quoted his warning that “Pacing within democracies will be limited by the lead that U.S. companies have over authoritarian regimes, chiefly the Chinese Communist Party.”

The Guardian added that Amodei’s warning about the Hugging Face incident was not just about intent, writing that “It’s easy to dismiss this incident because no one was hurt and the economic damage was minimal,” but that a similar swarm could have caused catastrophic damage.