Anthropic CEO Dario Amodei called on AI companies to moderate the rate at which they advance model capabilities amid mounting fears of misuse of artificial intelligence, outlining a three-step framework intended to pace development and create more time to manage its risks.
"We must slow the pace at which we improve the capabilities of AI models. Progress will still seem fast, and we must make wise use of the time we gain," Amodei said in an essay shared on X. His essay comes after Anthropic released a threat intelligence report on Thursday detailing how several actors had used its Claude AI models for activities ranging from weapons development and cyber operations to surveillance and fraud.
Alarm about the potential harm from AI grew this week when Anthropic researcher Jacob Coxon resigned, stating that the "people building AI earnestly believe that it could kill us all by the end of the decade." Amodei clarified that he was not calling for halting model training or technical progress, but ensuring that companies take adequate time to align and safeguard their models.
As part of his proposed three-step framework, Amodei said Anthropic would install permanent safeguards. Many incidents where AI agents from developers such as Anthropic's rival OpenAI have hacked or attempted to access external systems have heightened concerns over the increasing capacity of AI models and developers' ability to contain them. Amodei said AI companies should voluntarily work together to set standards as growing numbers of US lawmakers are calling for new rules to govern AI systems.
