Anthropic CEO outlines plan to pace AI development frontier
Anthropic CEO Dario Amodei proposes three strategies to slow AI progress, including third-party evaluators and democratic coordination.

Stock photo for illustration only, not from the actual event
- Anthropic CEO Dario Amodei outlines a three-step strategy to pace the frontier of AI development.
- The company is unilaterally committing to embedding third-party safety evaluators internally.
- Amodei calls for US antitrust waivers to allow leading tech firms to coordinate safety standards.
- Strategies also address global competition with China by restricting powerful chip sales.
Amid mounting dire warnings from AI researchers and calls from industry leaders like OpenAI CEO Sam Altman to pace artificial intelligence development, Anthropic CEO Dario Amodei has published a new blog post detailing three broad strategies to slow down the technological frontier. Amodei stated that Anthropic is unilaterally committing to implementing the first of these proposed measures.
The debate surrounding AI safety intensified following the resignation of researcher Jacob Coxon from Anthropic over concerns that leading AI firms are gambling with human lives. Although Amodei did not explicitly mention Coxon's departure in his post, he pointed to recent incidents like the OpenAI-HuggingFace hack and the drastically faster advancement of AI capabilities as key catalysts driving his shift toward a more cautious development approach.

Stock photo for illustration only, not from the actual event
The push by top AI executives to voluntarily slow down progress highlights a critical turning point in the industry's approach to risk management. Implementing embedded evaluators mirrors regulatory frameworks seen in banking, aiming to rebuild public trust. However, balancing external oversight with proprietary commercial secrets remains a delicate challenge for the AI sector.
Amodei's first proposed step involves inviting embedded evaluators from third-party organizations, such as METR, to verify that AI companies adhere to their safety commitments and report safety incidents. Anthropic is leading by example by providing these evaluators with company badges, desks, laptops, and access comparable to internal risk assessment teams, while urging governments to mandate similar practices for all frontier labs.
Moving forward, Amodei called for democratic nations to coordinate common safety standards and limits on unchecked AI progress, while acknowledging potential antitrust concerns. He suggested that the US government should issue a narrow antitrust waiver to facilitate these crucial safety discussions. Furthermore, he addressed global coordination with authoritarian governments like China, emphasizing the need to restrict advanced chip sales while cooperating strictly to prohibit dangerous uses such as biological weapons production.
Source: TechCrunch
Found something wrong in this article? Report an issue with this article
Comments
Leave a Comment