Anthropic CEO Calls for Slower Frontier AI Progress
Anthropic CEO Dario Amodei wrote in an essay that AI companies should slow the rate of capability advances so safety work can keep pace. He said Anthropic will give external evaluators ongoing access to review its safety practices, assess training processes and report incidents.
The evaluators will receive company laptops, office access and permissions similar to those held by internal risk teams, subject to legal, contractual and privacy limits. They will be able to publish findings without Anthropic's editorial control, although the company may redact protected or sensitive information.
Amodei also proposed common safety standards and capability limits among AI companies in democratic countries, supported by government regulation or coordination. He said requirements could link particular model capabilities to evaluations, interpretability reviews and audits of training environments.
His third step calls for international agreements covering dangerous AI uses, model testing and the pace of systems that help build more capable AI. Amodei cited faster recursive improvement and the OpenAI-Hugging Face cybersecurity incident as reasons for the proposal.
Amodei wrote that a swarm with greater capabilities but a similar level of misalignment could have caused catastrophic damage, and that his worry is that within 6 to 12 months such a swarm could be capable of taking over the entire internet with a persistent botnet, potentially causing hundreds of billions of dollars in damage. He named METR as an example of the kind of team Anthropic would embed, and said recursive self improvement is already starting to happen across the industry, including at Anthropic.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like AI Policy Brief or Daily AI Brief.
Also, consider following us on social media:
More from AI Safety
Sep 13 Sam Altman Tells OpenAI Staff It Could Slow AI Development Sep 13 Anthropic Blocks Yemen Cell Using Claude Code for Missile Software Sep 11 Anthropic Publishes Five Cases of Claude Use That Could Support Biological Weapons Work Sep 11 Senate Opens Inquiry Into OpenAI Agents' Hugging Face Hack Sep 11 Anthropic's Threat Report Finds AI Moving From Assistant to OrchestratorSubscribe to AI Policy Brief
Weekly report on AI regulations, safety standards, government policies, and compliance requirements worldwide.
Whitepaper
Tensordyne Napier: What If One Rack Could Do the Work of Nine?
Tensordyne
This Tensordyne whitepaper presents Napier, an inference-focused AI processor and rack-scale system based on the company’s TDN Math logarithmic number system. It examines infrastructure requirements for large mixture-of-experts and agentic models, compares major inference architecture approaches, and details the TDN AIP processor, TDN72 pod, TDN Link fabric, and Napier Ultra configuration. The paper reports simulation-based performance, cost, and accuracy-validation results, including Tensordyne’s projected comparison of one Napier rack with a nine-rack Nvidia Rubin plus Groq deployment; the chip is reported as taped out and in fabrication.
Read moreYou may also like
Sam Altman Tells OpenAI Staff It Could Slow AI Development
OpenAI Executive Warns of Persistent AI Cyber Attacks
Anthropic's Threat Report Finds AI Moving From Assistant to Orchestrator
Anthropic Researcher Jacob Coxon Resigns Over AI Safety Fears
OpenAI Says It Reached Its Automated Research Intern Goal
Daily AI Brief: the AI news that matters, in your inbox.