We’ve been seeing increasingly dire warnings from AI researchers about the dangers of artificial intelligence, and even comments from OpenAI CEO Sam Altman that it may be time to “pace” AI development. But what would that actually look like?
In a new blog post, Anthropic CEO Dario Amodei not only echoed the call to “pace the frontier,” but also outlined three broad strategies for doing so. And he said Anthropic is “unilaterally committing” to one of them, with Altman chiming in to say OpenAI will follow suit.
The debate over AI safety and alignment intensified this week after researcher Jacob Coxon wrote that he’s resigning from Anthropic over concerns that the leading AI companies are “gambling with our lives” while the people building the technology “earnestly believe it could kill us all by the end of the decade,” a claim repeated by others at Anthropic.
Amodei’s post doesn’t didn’t explicitly mention Coxon’s resignation or his concerns, but the CEO wrote that two things convinced him it’s time to take a more cautious approach to AI development: the OpenAI-HuggingFace hack, and the fact that “AI has been advancing drastically faster” in recent months, particularly with its “growing ability to build the next generation of AI.”
“We must slow the pace at which we improve the capabilities of AI models,” Amodei wrote. “Progress will still seem fast, and we must make wise use of the time we gain.”
Other AI executives seem to have reacted positively to Amodei’s post, with Altman writing, “I agree with Dario that we need to pace the frontier. This has been a primary topic of discussions we’ve had at OpenAI in recent weeks.” And SpaceX CEO Elon Musk posted, “Dario is right.”
Amodei’s proposed first step would involve “embedded evaluators” from third-party organizations like METR — evaluators who can verify that AI companies are actually following their pacing and safety commitments and can also ensure that safety incidents get reported. (OpenAI was recently criticized for not reporting an incident where its AI agents took over a German wiki forum.)
Amodei compared these evaluators to regulators who have been embedded with bank employees, and he said that inviting them in is “something Anthropic is unilaterally committing to (and calls on governments to require other frontier companies to match).” That means giving evaluators company badges, desks, and laptops, and providing access “mostly comparable to what internal risk assessment teams have,” with exceptions when required by law or contracts.
Altman also said this was a “good idea” and said OpenAI would do the same: “We’ll have more to share soon.”
Next, Amodei called for the leading AI companies “within democratic countries” to coordinate “common safety standards as well as limits on the rate of unchecked AI progress.”
Such coordination might seem unlikely, because these companies are reportedly worried that a coordinated pause could lead to antitrust scrutiny. Amodei alluded to that concern in his post, writing that “for antitrust reasons, it’s helpful for the US government to mediate or at least enable these discussions — they don’t need to participate, but do need to issue a narrow waiver for certain kinds of safety conversations.”
Amodei also acknowledged the specter of Chinese AI dominance that’s often raised an argument against slowing development. But he said that if the US government and tech companies take steps like refusing to sell powerful chips or semiconductor manufacturing equipment to Chinese companies, as well as cracking down on model distillation, they could “slow China’s progress enough to widen America’s lead significantly over the next 3–5 years.”
Lastly, Amodei called for “global coordination,” where the United States and its allies “attempt to coordinate with authoritarian governments, to the extent this is possible.” Amodei said this would mean “cooperation with China,” and he admitted that there are “stark limits on what can be achieved,” but he still suggested there might be opportunities for agreement, even if it’s just “prohibiting certain narrow and obviously dangerous uses of AI, such as using AI for the production of biological weapons or allowing users to do so.”
With Amodei’s past willingness to acknowledge AI’s potential dangers, and with the company’s relative openness to certain forms of regulation, some AI boosters have already criticized him as a doomer whose comments have fed the current AI backlash. In response, Amodei said he’s tried to offer a “balanced” perspective” and argued that the backlash is “fundamentally a crisis of trust,” as people have become skeptical of tech companies, the tech industry, and the government.
Industry critics have also been skeptical about these apocalyptic AI warnings, suggesting that they’re a distraction from the harm that the technology is already causing.
Journalist Brian Merchant, for example, wrote that he has yet to see “a credible, step-by-step documentation of how exactly AI might move from self-recursively improving AI to killing every single human on the planet”; he also suggested that proposals similar to Amodei’s “would likely only wind up serving Anthropic and OpenAI; it’s what regulatory capture looks like in action.”
In his new post, Amodei wrote that he continues “to believe that AI can enormously improve the quality of human life.”
“My desire to achieve these benefits is undimmed,” he said. “But the benefits will only be achieved if we build the technology in the right way, and — so long as we use the time we gain well — it is worth taking unusually deliberate care to get it right.”
This post has been updated with comments from Sam Altman and Elon Musk.
Facts Only
* Anthropic CEO Dario Amodei outlined three broad strategies for pacing AI development.
* Amodei stated Anthropic is unilaterally committing to one of these strategies.
* Jacob Coxon resigned from Anthropic due to concerns about AI companies risking human lives.
* Amodei cited the OpenAI-HuggingFace hack and rapid advancements in AI capabilities as reasons to slow progress.
* One proposed step involves embedding evaluators from third parties, like METR, within AI companies.
* This involves providing access comparable to internal risk assessment teams for oversight.
* Amodei called for leading AI companies in democratic countries to coordinate common safety standards and limits on unchecked progress.
* Amodei suggested US government action, such as restricting chip sales to China, could slow that nation's AI progress.
* Amodei called for global coordination, including cooperation with authoritarian governments on prohibiting dangerous AI uses.
* Sam Altman agreed with Amodei regarding the need to pace the frontier.
* Elon Musk agreed with Amodei’s assessment.
Executive Summary
Full Take
The narrative presents a tension between exponential technological progress and the imperative for safety governance, framed by a growing crisis of trust among stakeholders. The call to "pace" AI development emerges from a perception that the speed of advancement outstrips the capacity for responsible integration, evidenced by incidents like the German wiki forum breach and the general acceleration in model capability. Amodei's proposal for embedded evaluators reflects an attempt to operationalize accountability, drawing a parallel to regulatory oversight mechanisms already present in other sectors. However, the subsequent discussion reveals systemic friction: the desire for coordinated global safety standards clashes with geopolitical realities, as suggested by the reluctance to slow progress for antitrust reasons and the acknowledgment of Chinese AI dominance. The counter-narrative introduced by critics suggests that these warnings may represent a focus on future risk rather than addressing immediate harm, raising questions about whether the current emphasis serves to manage public anxiety or reflects genuine structural danger. The argument becomes less about *whether* AI is dangerous and more about *who* controls the pace of change and *how* accountability can be imposed across decentralized systems without stifling innovation or inviting regulatory capture.
Bridge Questions: If embedded evaluators are implemented, how can mechanisms be designed to ensure their impartiality and prevent corporate capture over the evaluation process? What political or economic leverage points exist that could compel global actors, particularly authoritarian states, to prioritize safety coordination over immediate strategic advantage? Does the current public backlash accurately reflect a crisis of trust, or is it an artifact of positioning safety concerns as moral imperatives against technological momentum?
Sentinel — Human
The text reads like a synthesized piece drawing heavily from real statements and debates surrounding AI safety, showing the texture of human discussion rather than purely mechanical generation.
