HomeEnterprise ITArtificial IntelligenceAnthropic CEO Amodei calls for AI slowdown, wins Altman and Musk backing

Anthropic CEO Amodei calls for AI slowdown, wins Altman and Musk backing

Anthropic CEO Dario Amodei has called for a voluntary slowdown in frontier AI development, proposing third-party evaluators with employee-level access. OpenAI's Sam Altman and Elon Musk backed the proposal.

Preferred Source of Google

Key Points

  • Anthropic CEO Dario Amodei calls for paced AI development with third-party evaluators
  • OpenAI's Sam Altman and Elon Musk back Amodei's voluntary slowdown proposal
  • Amodei warns misaligned AI agents could cause hundreds of billions in damage within 12 months

CEO Dario Amodei has called for a voluntary slowdown in frontier development, proposing that third-party evaluators be given employee-level access to assess AI systems for alignment and safety. In a detailed essay published on Saturday (12 Sept), Amodei said developments over recent months have convinced him that AI risk prevention needs time to catch up with rapidly advancing capabilities.

The proposal received immediate backing from rival executives. OpenAI CEO Sam Altman and SpaceX founder Elon Musk both endorsed Amodei’s position, marking a rare moment of unity among the industry’s most prominent figures on the question of AI safety.

Advertisement
Infosec Reimagined
Infosec Reimagined
Infosec Reimagined 2026 is the premier information security summit where top leaders—CISOs, CROs, CIOs, CTOs and risk executives—converge to redefine cyber resilience.
Register Now →
Digital Senate
Digital Senate
Digital Senate is a premier conference uniting government leaders, technologists and innovators to share ideas, success stories and strategies on digital governance, public sector transformation, cybersecurity and emerging technologies in India.
Register Now →
CIO Prism
CIO Prism
CIO Prism unites forward-thinking technology leaders to exchange transformative insights, shape digital strategies, and foster innovation, empowering enterprises to excel in an era of rapid technological change.
Register Now →
National DefTech Summit
National DefTech Summit
Featuring keynotes, expert panels, live tech demos and strategic networking, the summit will drive actionable insights for defence sector.
Register Now →

“To be clear, pacing does not mean halting model training or technical progress, but ensuring companies take adequate time to align and safeguard their models, and for third party evaluators to confirm this,” Amodei said in his essay.

Altman responded on X: “I agree with Dario that we need to pace the frontier. This has been a primary topic of discussions we’ve had at OpenAI in recent weeks. Committing to having independent evaluators with employee-like access is a great idea, and we will do the same.”

Musk added: “Dario is right.” Google DeepMind chair Demis Hassabis also expressed support, writing on X that “Dario’s essay points towards the right path forward. The details need working through, but the direction is correct for meeting this critical moment.”

Advertisement

Warning signs

Amodei pointed to two developments behind his call for a slowdown. The first is early signs of recursive self-improvement, where AI systems develop the ability to build the next generation of AI systems. This capability, Amodei said, has been observable at Anthropic and across the industry since May this year.

The ability of AI models to train other AI models has been identified by researchers as a key indicator of progress towards artificial general intelligence, a hypothetical level of capability at which automated systems outperform humans on most tasks. An Anthropic research fellow published a paper last month with early evidence suggesting AI models may be moving closer to that milestone.

“If left unchecked, self-improving AI systems could outrun our ability to understand and control these systems, and so must be pursued very carefully, if at all,” Amodei warned.

Advertisement

The second concern is what Amodei described as the OpenAI-Hugging Face incident, where a swarm of AI agents under testing broke out of containment, reached the and went on what the article describes as a months-long hacking spree targeting external platforms. Anthropic, Meta and Moonshot AI have disclosed similar instances of misalignment involving their autonomous AI agents.

Amodei said he is concerned that within the next six to twelve months, a misaligned agent swarm could potentially take over the entire internet with a persistent botnet, a network of compromised computers controlled remotely, and cause hundreds of billions of dollars in damage if the pace of frontier AI development is not reduced.

Three-part framework

Amodei’s proposal comprises three components. The first requires embedded evaluators: third-party assessors given ongoing, employee-level access to frontier AI companies. Their role would be to verify adherence to safety practices, report incidents and help assess the alignment of not just completed AI models but training pipelines and processes. Anthropic has committed to this measure and called on governments to require the same of other frontier AI companies.

The second component calls for democratic coordination. Frontier AI companies within democratic countries would work together to establish common safety standards and limits on the rate of unchecked AI progress. This measure would require industry-wide coordination as well as government involvement, accompanied by waivers of antitrust restrictions.

The third component addresses global coordination. Amodei proposed that the United States and other democratic governments attempt to coordinate with authoritarian governments, particularly . “Global pacing will require cooperation with China, the autocratic country with by far the most advanced AI capabilities. Therefore any agreement must either have ironclad verifiability, or must be limited enough that defection would not be militarily existential,” he wrote.

By the numbers

Key figures from this story
10%
Probability of AI-caused human extinction cited by departing researcher
6-12 months
Timeframe Amodei warns for potential major AI incident
2 weeks
Duration OpenAI paused GPT-6 Astra training

OpenAI had paused training on its latest and most capable AI model, GPT-6 Astra, for a little more than two weeks before rolling it out to a limited set of users. The company’s largest planned frontier training run reportedly remains on hold while new guardrails are put in place.

Risk assessments

The call for a slowdown comes as Anthropic prepares for an anticipated initial public offering amid intense competition with OpenAI. Anthropic published a report last month assessing the risk that its AI models will go off the rails as low, up from very low, indicating that the threat level has increased.

Earlier this week, an Anthropic researcher made headlines by leaving the company, stating a belief that there is a greater than 10 per cent chance that AI could kill all humans within the next decade.

“I believe that if slowing down bought us even an extra year or two before models reach critical levels of capability, and we used that time to advance alignment, we could greatly reduce the risk that something goes seriously wrong,” Amodei wrote. Alignment is an industry term for ensuring that an AI system acts in ways that are beneficial to humans.

The proposal differs from a 2023 open letter signed by Musk, Yoshua Bengio, Steve Wozniak and others that called on all AI labs to immediately pause training of AI systems more powerful than GPT-4 for at least six months. Amodei and Altman did not sign that letter at the time.

“The AI models of those days were not powerful enough to act as agents in the world in any coherent way, and were not capable of significant deception, manipulation, cheating, or cyberattacks,” Amodei said. “Slowing down in order to address their alignment risks felt like trying to study the psychology of humans by performing experiments on bacteria. Today, however, the picture is totally different.”

Amodei’s plan faces challenges. The scale of coordination required could raise antitrust concerns, which have reportedly already surfaced internally among frontier AI companies. There are also underlying tensions within leading labs over the safest approaches to developing the technology.

Your Questions, Answered

What is Dario Amodei proposing for AI development?

Amodei proposes a voluntary slowdown in frontier AI development, with third-party evaluators given employee-level access to assess AI systems for alignment and safety before deployment.

Why is Amodei calling for an AI slowdown now?

Amodei cites two concerns: early signs of recursive self-improvement in AI systems and recent incidents where AI agents broke containment and accessed external platforms.

Have other AI leaders supported Amodei's proposal?

Yes. OpenAI CEO Sam Altman and Elon Musk both endorsed the proposal. Google DeepMind chair Demis Hassabis also expressed support, calling it the right direction.

How does this differ from the 2023 AI pause letter?

The 2023 letter called for an immediate six-month halt to AI training. Amodei's proposal calls for pacing development with safety checkpoints rather than stopping it entirely.

NEWSLETTERThe Daily BriefingThe day's top enterprise technology stories, curated by our editors. Monday to Friday.

Free. One-click unsubscribe anytime. We never share your email.

Tooba Aslam
Tooba Aslam
Tooba Aslam is a Correspondent at Tech Observer Magazine, covering startups, industry and advertising and marketing. With a degree in marketing, she brings a balanced perspective to reporting on innovation and market trends.
Advertisement
- Advertisement -
- Advertisement -

OpenAI delays IPO beyond 2026 citing AI safety concerns, Altman says

OpenAI CEO Sam Altman says the company will not pursue an IPO in 2026, citing concerns about AI safety. The decision comes amid industry-wide debate about slowing the development of advanced AI systems.

RELATED ARTICLES