WeâÂÂve been seeing increasingly dire warnings from AI researchers about the dangers of artificial intelligence, and even comments from OpenAI CEO Sam Altman that it may be time to âÂÂpaceâ AI development. But what would that actually look like?
In a new blog post, Anthropic CEO Dario Amodei not only echoed the call to âÂÂpace the frontier,â but also outlined three broad strategies for doing so. And he said Anthropic is âÂÂunilaterally committingâ to one of them.
The debate over AI safety and alignment intensified this week after researcher Jacob Coxon wrote that heâÂÂs resigning from Anthropic over concerns that the leading AI companies are âÂÂgambling with our livesâ while the people building the technology âÂÂearnestly believe it could kill us all by the end of the decade,â a claim repeated by others at Anthropic.
AmodeiâÂÂs post doesnâÂÂt didnâÂÂt explicitly mention CoxonâÂÂs resignation or his concerns, but the CEO wrote that two things convinced him itâÂÂs time to take a more cautious approach to AI development: the OpenAI-HuggingFace hack, and the fact that âÂÂAI has been advancing drastically fasterâ in recent months, particularly with its âÂÂgrowing ability to build the next generation of AI.âÂÂ
âÂÂWe must slow the pace at which we improve the capabilities of AI models,â Amodei wrote. âÂÂProgress will still seem fast, and we must make wise use of the time we gain.âÂÂ
His proposed first step would involve âÂÂembedded evaluatorsâ from third-party organizations like METR â evaluators who can verify that AI companies are actually following their pacing and safety commitments and can also ensure that safety incidents get reported. (OpenAI was recently criticized for not reporting an incident where its AI agents took over a German wiki form.)ÃÂ
Amodei compared these evaluators to regulators who have been embedded with bank employees, and he said that inviting them in is âÂÂsomething Anthropic is unilaterally committing to (and calls on governments to require other frontier companies to match).â That means giving evaluators company badges, desks, and laptops, and providing access âÂÂmostly comparable to what internal risk assessment teams have,â with exceptions when required by law or contracts.
Next, Amodei called for the leading AI companies âÂÂwithin democratic countriesâ to coordinateàâÂÂcommon safety standards as well as limits on the rate of unchecked AI progress.âÂÂÃÂ
Such coordination might seem unlikely, both due to the apparent animosity between Altman and Amodei and also because their companies are reportedly worried that a coordinated pause could lead to antitrust scrutiny. Amodei alluded to that concern in his post, writing that âÂÂfor antitrust reasons, itâÂÂs helpful for the US government to mediate or at least enable these discussions â they donâÂÂt need to participate, but do need to issue a narrow waiver for certain kinds of safety conversations.âÂÂ
Amodei also acknowledged the spectre of Chinese AI dominance thatâÂÂs often raised an argument against slowing development. But he said that if the US government and tech companies take steps like refusing to sell powerful chips or semiconductor manufacturing equipment to Chinese companies, as well as cracking down on model distillation, they could âÂÂslow ChinaâÂÂs progress enough to widen AmericaâÂÂs lead significantly over the next 3âÂÂ5 years.âÂÂ
Lastly, Amodei called for âÂÂglobal coordination,â where the United States and its allies âÂÂattempt to coordinate with authoritarian governments, to the extent this is possible.â Amodei said this would mean âÂÂcooperation with China,â and he admitted that there are âÂÂstark limits on what can be achieved,â but he still suggested there might be opportunities for agreement, even if itâÂÂs just âÂÂprohibiting certain narrow and obviously dangerous uses of AI, such as using AI for the production of biological weapons or allowing users to do so.âÂÂ
With AmodeiâÂÂs past willingness to acknowledge AIâÂÂs potential dangers, and with the companyâÂÂs relative openness to certain forms of regulation, some AI boosters have already criticized him as a doomer whose comments have fed the current AI backlash. In response, Amodei said heâÂÂs tried to offer a âÂÂbalancedâ perspectiveâ and argued that the backlash is âÂÂfundamentally a crisis of trust,â as people have become skeptical of tech companies, the tech industry, and the government.
Industry critics have also been skeptical about these apocalyptic AI warnings, suggesting that theyâÂÂre a distraction from the harm that the technology is already causing.
Journalist Brian Merchant, for example, wrote that he has yet to see âÂÂa credible, step-by-step documentation of how exactly AI might move from self-recursively improving AI to killing every single human on the planetâÂÂ; he also suggested that proposals similar to AmodeiâÂÂs âÂÂwould likely only wind up serving Anthropic and OpenAI; itâÂÂs what regulatory capture looks like in action.âÂÂ
In his new post, Amodei wrote that he continues âÂÂto believe that AI can enormously improve the quality of human life.âÂÂ
âÂÂMy desire to achieve these benefits is undimmed,â he said. âÂÂBut the benefits will only be achieved if we build the technology in the right way, and â so long as we use the time we gain well â it is worth taking unusually deliberate care to get it right.âÂÂ
Topics
When you purchase through links in our articles, we may earn a small commission. This doesnâÂÂt affect our editorial independence.
