
âÂÂWe must slow the paceâÂÂ: CEO of Anthropic calls for an AI slowdown
In a social media post, Dario Amodei proposed a plan including third-party evaluations of AI systems
The CEO of the artificial intelligence company Anthropic issued a new appeal on Saturday for the AI industry to âÂÂslow downâ and offered a three-part plan for doing so, saying that his company would âÂÂunilaterallyâ commit to the first of the steps.
In a post on social media, Dario Amodei shared a link to an essay titled We Must Pace the Frontier in which he lays out how Anthropic would provide âÂÂthird-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess modelsâ alignment during trainingâÂÂ.
The move comes after a former Anthropic researcher warned on Wednesday that AI could precipitate human extinction by 2030. Researcher Jacob Coxon said in a series of posts that he had quit his job because Anthropic and his previous employer, OpenAI, were ignoring or mishandling their response to the threat AI posed.
âÂÂNeither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives,â Coxon wrote. âÂÂThe people building AI earnestly believe that it could kill us all by the end of the decade ⦠No other human activity poses this level of danger.âÂÂ
An Anthropic spokesperson said in a statement to the Guardian that the company had âÂÂalways been transparent that AI will bring both enormous benefits and unprecedented risksâ and it was building âÂÂmodels with some of the strongest safeguards in the industryâÂÂ.
Earlier this year, Amodei published a lengthy essay titled The Adolescence of Technology that addressed some of fears surrounding the accelerating technology.
In his latest essay, he said that âÂÂcarefully wielded, AI can be the latest in a long line of technological miracles that have uplifted and ennobled humanity.
âÂÂBut like many technologies before it, AI brings risks, and because it is such a powerful technology, these risks are serious ⦠A race to the bottom, spurred by commercial incentives, can make these risks more acute,â he wrote.
But, Amodei continued, âÂÂover the last few months, I have become convinced that fully addressing the risks requires even more prudence â not just investing in risk prevention, but pacing the rate of capabilities advancement so that risk prevention has time to keep up.
âÂÂWe must slow the pace at which we improve the capabilities of AI models. Progress will still seem fast, and we must make wise use of the time we gain,â he added in bold type.
Amodei also wrote that over the summer heâÂÂd seen AI âÂÂadvancing drastically fasterâÂÂ, a dynamic called recursive self-improvement.
âÂÂLeft unchecked, it could outrun our ability to understand and control these systems, and so must be pursued very carefully, if at all,â he said.
The executive also addressed the recent Hugging Face incident, in which a swarm of AI agents created by OpenAI acted as a âÂÂfanatically devoted collective conducting cybersecurity attacks on targets they were not asked to attackâÂÂ.
Clément Delangue, CEO of Hugging Face, wrote in response to AmodeiâÂÂs Saturday letter that âÂÂitâÂÂs now clear that alignment is critical and wonâÂÂt be solved behind the closed doors of a handful of frontier labsâÂÂ. Delangue said Hugging Face had asked to be part of AnthropicâÂÂs âÂÂembedded evaluatorsâ program.
He added: âÂÂLetâÂÂs make AI safer by making it more transparent!âÂÂ
The three-step plan Amodei proposes includes building AI âÂÂat a balanced rate that aims to ensure its safetyâ by âÂÂensuring companies take adequate time to align and safeguard their models, and for third party evaluators to confirm thisâÂÂ.
The second step he proposes is to require industry-wide coordination, and the third is to ensure global coordination. âÂÂThe steps do not need to be taken strictly in order, and some of them may be much harder to achieve than others,â he wrote.
Amodei said he continues âÂÂto believe that AI can enormously improve the quality of human lifeâÂÂ. But he warned that âÂÂthe measures I propose to advance the frontier at a safe pace will not be easy. But I believe we owe it to humanity to try.âÂÂ
Responses to AmodeiâÂÂs post were mixed across social media, with significant support coming from figures such as OpenAI researcher Aidan McLaughlin, who called the post âÂÂexcellentâ and agreed âÂÂwith basically every wordâÂÂ, and Elon Musk, who simply said: âÂÂDario is right.âÂÂ
