Anthropic’s Dario Amodei Calls for Slower AI Progress and Outside Safety Reviews
The Anthropic chief is pairing an immediate transparency commitment with a harder request: competing labs and governments should agree to constrain a race he says could outrun human control.
Listen to this story
The audio brief
Story brief
3 key pointsAnthropic is pairing a call for slower frontier-model capability gains with an immediate governance commitment: independent safety evaluators will receive employee-level access and may publish findings without company editorial control. CEO Dario Amodei also wants rival labs and governments to adopt permanent evaluations, shared standards among democratic countries, and agreements covering risks such as AI-assisted...
- 01
The evaluator arrangement takes effect immediately and is intended to match access given to Anthropic’s internal risk-assessment teams.
- 02
Amodei warns recursive self-improvement could accelerate development beyond human understanding or control.
- 03
His proposal covers common safety standards among democratic countries and coordination with authoritarian states on shared risks.
Anthropic CEO Dario Amodei is urging frontier AI developers to slow capability gains while making one immediate concession of his own: permanent access for independent safety evaluators inside Anthropic. His proposal puts a practical question at the center of the AI race: can a company’s voluntary openness become a basis for collective restraint among competitors and governments?
In an essay posted to his website, Amodei wrote that the pace of improvement in AI models should slow, even as progress remains fast, so developers can use the time to improve safety and judgment. He warned that misuse or loss of control could bring cyberattacks, bioterrorism and serious economic disruption.
One company can open its doors; a race needs shared rules
Anthropic says its independent evaluators will have employee-level access to examine safety practices and report incidents. They will receive the same access as the company’s internal risk-assessment teams and may publish findings without Anthropic’s editorial control. Amodei said the arrangement takes effect immediately.
We must slow the pace at which we improve the capabilities of AI models. Progress will still seem fast, and we must make wise use of the time we gain.
Dario Amodei, in his essay
That is the part of the agenda under Anthropic’s direct control. The rest is a request for alignment beyond the company: Amodei proposed permanent independent evaluation across frontier AI companies, common safety standards among democratic countries, and coordination between democratic governments and authoritarian states on shared interests, including a ban on using AI to develop biological weapons.
The rationale for a slower pace
Amodei’s case is not that AI progress should stop. It is that increasingly capable systems may help create their successors, accelerating development faster than people can understand or control it. He described unchecked recursive self-improvement as a dynamic that must be pursued very carefully, if at all.
The proposal has three connected parts
- Make independent evaluation a permanent feature of frontier AI companies.
- Set common safety standards among democratic countries to limit unchecked progress.
- Seek international agreements on risks with mutual stakes, including AI-enabled biological weapons development.
A commitment at Anthropic, a proposal beyond it
Anthropic has committed to a standing review mechanism with employee-level access and publication rights for independent evaluators. The wider objective—slowing capability gains while preventing dangerous uses—depends on whether other companies and governments accept Amodei’s proposals for shared evaluation, common standards and international coordination.
Editorial analysis
Our Read
Anthropic’s immediate commitment is narrow but concrete: outsiders are meant to gain durable access and publication rights inside one lab. The larger part of Amodei’s plan depends on competitors and governments accepting common limits, which no single company can impose. That makes the next meaningful test less about the essay’s warning than whether other frontier developers adopt comparable scrutiny or shared thresholds for slowing work. The proposal also arrives after other public calls for coordinated restraint, underscoring that agreement—not simply concern—is the unresolved operating problem.
Citation desk / original work
Cite this
Citation desk / original work
Cite this
Anthropic’s immediate commitment is narrow but concrete: outsiders are meant to gain durable access and publication rights inside one lab.
/posts/anthropic-s-dario-amodei-calls-for-slower-ai-progress-and-outside-safety-reviews#finding-1
Sources
- fortune.comAnthropic CEO calls to slow the race toward AI 'superintelligence' | Fortune
- kwwl.comAnthropic CEO calls for ‘pacing the frontier’ of AI race amid safety concerns
- kten.comAnthropic CEO calls for ‘pacing the frontier’ of AI race amid safety concerns
Loading discussion...
Reader comments
Newest comments first. Replies stay oldest first.