
Warnings from AI researchers regarding the perils of artificial intelligence have become increasingly urgent, with OpenAI CEO Sam Altman suggesting it may be time to “pace” AI progress. But what would that entail in practice?
In a recent blog entry, Anthropic CEO Dario Amodei not only supported the idea of “pacing the frontier,” but also proposed three overarching strategies for achieving it. He mentioned that Anthropic is “unilaterally committing” to one of these strategies.
The discussion surrounding AI safety and alignment escalated this week after researcher Jacob Coxon stated that he is leaving Anthropic because he believes that leading AI firms are “gambling with our lives,” while those who develop the technology “sincerely believe it could endanger humanity by the decade’s end,” a sentiment echoed by others at Anthropic.
Although Amodei’s post did not directly reference Coxon’s resignation or his worries, the CEO expressed that two events persuaded him to adopt a more cautious stance on AI development: the OpenAI-HuggingFace breach, and the observation that “AI has been advancing at an unprecedented pace” recently, especially regarding its “increasing capability to create the next generation of AI.”
“We need to decelerate the rate at which we enhance the abilities of AI models,” Amodei asserted. “Progress will still appear rapid, and we need to wisely utilize the additional time we create.”
His initial proposal would incorporate “embedded evaluators” from independent organizations like METR — evaluators tasked with verifying that AI firms adhere to their pacing and safety pledges while also ensuring that safety issues are reported. (OpenAI faced criticism for not disclosing an incident involving its AI agents taking control of a German wiki form.)
Amodei likened these evaluators to regulators who have been integrated with banking personnel, and stated that engaging them is “something Anthropic is unilaterally committing to (and encourages governments to mandate other leading companies to comply).” This entails providing evaluators with corporate badges, workspaces, and laptops, and granting them access that is “generally comparable to that which internal risk assessment teams have,” with legal or contractual exceptions.
Following that, Amodei urged the principal AI companies “in democratic nations” to collaborate on “shared safety standards as well as restrictions on the pace of unregulated AI advancement.”
This kind of cooperation might seem improbable, given the clear tension between Altman and Amodei, as well as the fears among their companies that a unified suspension could draw antitrust scrutiny. Amodei hinted at these concerns in his article, stating that “for antitrust purposes, it would be beneficial for the US government to facilitate or at least support these conversations — they need not be participants, but must issue a limited waiver for specific types of safety discussions.”
Amodei also recognized the looming threat of Chinese AI supremacy frequently cited as an argument against slowing down development. Nonetheless, he asserted that if the US government and tech firms take measures like withholding powerful chips or semiconductor production equipment from Chinese businesses and restricting model distillation, they could “deter China’s progress sufficiently to significantly extend America’s lead in the next 3–5 years.”
Ultimately, Amodei advocated for “global coordination,” where the United States and its allies “attempt to engage with authoritarian regimes, as much as practicable.” Amodei specified this would entail “collaboration with China,” acknowledging that there are “clear limitations on what can be accomplished,” yet he proposed there could be possibilities for agreement, perhaps only “banning specific narrow and evidently hazardous AI applications, such as employing AI for creating biological weapons or permitting users to do so.”
Given Amodei’s previous acknowledgment of AI’s potential risks, and the company’s relative openness to some regulation, some advocates of AI have criticized him as a pessimistic figure whose remarks have contributed to the current AI backlash. In reply, Amodei indicated that he has sought to present a “balanced” viewpoint and contended that the backlash constitutes “a fundamental crisis of trust,” as the public has grown skeptical of technology firms, the tech sector, and government entities.
Critics from the industry have also expressed doubt regarding these doomsday AI warnings, asserting they serve as a diversion from the harm that the technology is already inflicting.
Journalist Brian Merchant, for instance, stated he has yet to observe “a credible, detailed account of how precisely AI might leap from self-improving AI to exterminating every human on Earth”; he also suggested that proposals akin to Amodei’s “would likely only end up benefiting Anthropic and OpenAI; it exemplifies regulatory capture in action.”
In his latest post, Amodei affirmed his belief that AI can significantly enhance human life quality.
“My aspiration to realize these benefits remains unwavering,” he remarked. “However, these advantages will only materialize if we develop the technology appropriately, and — as long as we effectively utilize the time we acquire — it is essential to take unusually careful measures to ensure we get it right.”
When you make purchases through links in our articles, we might earn a small commission. This does not compromise our editorial autonomy.

