News & Blog

Anthropic boss Dario Amodei calls for AI development to slow down

The head of AI company Anthropic has called for the pace of development of artificial intelligence models to slow down and to be closely monitored.

Dario Amodei wrote in an online essay that developing AI was not in question, but that the risks associated with it were “serious” and that companies and governments must be given time to address them.

The bosses of two rival AI firms, Sam Altman of OpenAI and Elon Musk, have both said they agree with Amodei.

There have been growing concerns recently about the technology’s potential risks. An AI researcher who left Anthropic told the BBC that “if we don’t slow down at the current rate of progress, there is a strong chance that we could all die in the immediate future”.

Jacob Coxon told Laura Kuenssberg that people working at AI companies were “genuinely frightened… genuinely concerned about the fate of humanity in the next two years”.

“It’s not at all an exaggeration to say that the people who are involved with both founding these companies and building the tech believe there is a possibility of human extinction,” Coxon said.

In Amodei’s essay, called We Must Pace the Frontier and published on Saturday, he proposed a three-point plan that included independent monitoring of AI models as they are developed, industry-wide regulation and global regulation.

As his proposal made the rounds, even competitors voiced support for the idea of third-party monitors who could evaluate the safety of models as they are developed.

“I agree with Dario that we need to pace the frontier,” wrote OpenAI CEO Sam Altman on X. He called independent evaluators “a great idea”.

Altman sounded similar safety concerns in a new interview, telling Fortune magazine that standards were “not at a place” to push AI capabilities much further.

He added that he believed AI beyond human control was “absolutely” possible.

Musk, meanwhile, who founded Grok producer xAI, said the Anthropic boss was “right”.

The warnings have prompted calls to action, but US President Donald Trump has so far rejected such fears, saying on Thursday he was concerned that “if we don’t win AI, we’re going to be put in a very bad position”.

Cyber-security concerns have grown as new models have exhibited more and more powerful capabilities.

Anthropic withheld its Mythos model from public use when it was announced in April that it could independently escape the testing environment, known as the sandbox.

In the run-up to the release of its most recent Astra model, OpenAI cited cyber-security concerns as it explained it had paused certain aspects of the model’s development.

The OpenAI agents had “essentially acted as a fanatically devoted collective”, Amodei said. OpenAI has said it was slowing down training of certain advanced AI models and tools as a result.

Safety has also taken centre stage in the rivalry between Anthropic and OpenAI.

Amodei, who had previously worked as a vice-president at OpenAI, has said he co-founded Anthropic in 2021 so he could build safer and more trusted AI models.

He pointed out in his essay that AI had advanced “drastically faster” including its “ability to build the next generation of AI” – and mentioned an incident involving rival OpenAI which has revealed that agents had hacked, external targets they were not asked to attack in July.

Amodei called for “building AI at a balanced rate that aims to ensure its safety while still achieving its benefits”.

This would not mean “halting model training or technical progress, but ensuring companies take adequate time to align and safeguard their models, and for third party evaluators to confirm this”.

He was committing Anthropic to this “unilaterally” – as well as calling on governments “to require other frontier companies to match”.

Amodei said he recognised that regulation might not be able to keep up with the pace of AI, and therefore called on AI companies to “voluntarily work together to set standard” in parallel with those regulation.

The Anthropic CEO went on to address the impact that a slowdown would have on the industry and competition with leading developers worldwide, particularly China.

“I believe that if slowing down bought us even an extra year or two before models reach critical levels of capability, and we used that time to advance alignment, we could greatly reduce the risk that something goes seriously wrong,” Amodei said.

This would have to be done in a co-ordinated manner “without sacrificing commercial advantage or the United States’ lead in AI”. Any slowdown would have to be limited, he said, to avoid allowing China to pull ahead.

He urged the US government to take measures so that US companies’ AI chips could not be sold to China – or the technology shared with authoritarian countries.

His former employee, Jacob Coxon, told the BBC, however, that the initiative to slow down must go beyond the US – “there’ll need to be some sort of co-ordinated slowdown with China if we’re going to avoid a race at an international scale”.

Amodei’s post has prompted a wide range of responses.

Clement Delangue, the CEO of the AI platform Hugging Face, said he was launching a new project called the Open Alignment Initiative, adding that he wanted to be among “embedded evaluators” that Amodei proposed could be part of a solution.

Hugging Face was hacked by OpenAI agents earlier this year, prompting outcry over AI safety.

“Let’s make AI safer by making it more transparent,” Delangue wrote on X.

Musk, who also voices support for Amodei, once called Anthropic “evil” but has changed his tone since signing a $15bn (£11bn) deal to sell computing capacity to Anthropic in May.

However, some observers suggested that Amodei’s post was less about safety than about consolidating control over AI technology.

“Dario makes the case to stop open source and concentrate enormous technological and economic power with Anthropic,” wrote Chamath Palihapitiya, an investor and co-host of the tech podcast “All-In”.

Notions of slowing down or even pausing AI development have long been met with cynicism in certain corners of Silicon Valley, with critics accusing leading AI developers of hyping their technology as a marketing ploy.

Anthropic and OpenAI are both reportedly preparing for potentially record-setting initial public offerings.