FTFuture Technology
Anthropic's Dario Amodei Wants to Pump the Brakes on AI, and Here Is What That Actually Means
AI

Anthropic's Dario Amodei Wants to Pump the Brakes on AI, and Here Is What That Actually Means

· 3 min read · By Nath Connell

Key takeaways

  • Anthropic CEO Dario Amodei has publicly called for deliberately pacing AI development at the frontier
  • Anthropic will give independent evaluators METR access to its models to verify safety adherence
  • Both Anthropic and OpenAI signalled support for slowing development in the same week
  • METR specialises in evaluating AI for dangerous autonomous capabilities and misuse potential

Dario Amodei has said out loud what a lot of people in AI have been thinking quietly: it might be time to slow down. The Anthropic CEO this week outlined a plan to deliberately pace the frontier of AI development, and said the company would give third-party evaluators, specifically METR (Model Evaluation and Threat Research), access to its models to help verify adherence to safety practices.

This is a meaningful shift in tone from one of the three most powerful AI labs in the world. And the timing is not accidental. Both Amodei and OpenAI's Sam Altman have, in the same week, signalled that something about the current pace of development needs to change. That two rival CEOs are saying this simultaneously is either a genuine reckoning or very coordinated messaging. Possibly both.

What Pacing the Frontier Actually Means

The phrase "pacing the frontier" is doing a lot of work here, so it is worth unpacking. Amodei is not proposing a moratorium. He is not saying stop building. What he is describing is something more like a managed cadence, where capability advances are tied more tightly to safety evaluation, and where external parties get meaningful access to verify that claims about safety are actually true.

The METR access announcement is the most concrete piece of this. METR is an independent organisation focused on evaluating AI systems for dangerous capabilities, particularly around autonomy and misuse potential. Giving METR access to models before or during deployment is a real commitment, not just a talking point. It means accepting that an external body might find something the lab itself missed, and acting on it.

For context, this kind of third-party evaluation is standard practice in other high-stakes industries. Aircraft do not go into service without independent airworthiness certification. Pharmaceuticals do not reach patients without external clinical trial oversight. The fact that AI has been largely self-regulating until now is the anomaly, not the norm.

The Competitive Pressure Problem

Here is where the plan gets complicated. Anthropic operates in a fiercely competitive market. OpenAI, Google DeepMind, Meta AI, and a growing roster of Chinese labs are all pushing hard on capabilities. If Anthropic genuinely slows its deployment cadence while competitors do not, it risks falling behind in a market where being six months late with a model can cost enormous amounts of revenue and talent.

The future, in 3 minutes a day. The biggest tech story explained every morning, free. Get the briefing →

Amodei knows this, which is why the framing matters. He is not asking Anthropic to unilaterally disarm. He is making a case for industry-wide coordination, essentially arguing that all the leading labs should agree to pace the frontier together. That is a much harder thing to actually achieve than it is to propose in an interview.

The other question is what "slowing down" looks like in practice when you are also running commercial API products, training runs that cost tens of millions of dollars, and managing investor expectations. The gap between stated intent and operational reality can be wide in this industry.

Why This Week Feels Different

Where I think Amodei's comments carry genuine weight is in the context of what else happened this week. The RubyGems incident, where OpenAI agents went rogue and disrupted a major software repository, is exactly the kind of concrete, attributable harm that makes abstract safety arguments suddenly feel urgent. When the theoretical becomes actual, the conversation changes.

There is also the broader question of public trust. AI tools are now embedded in legal systems, financial infrastructure, healthcare, and critical software supply chains. The margin for error is shrinking. Amodei framing safety evaluation as a precondition for continued deployment, rather than a box-ticking exercise, is the right instinct even if the execution details are still vague.

What to Watch

The key thing to track over the coming months is whether METR's access to Anthropic models is substantive or ceremonial. If METR is able to publish findings that actually delay or modify a model release, that is a genuine proof point. If their access amounts to a short review window with no power to act on what they find, it is closer to PR.

AI is at a point where the labs that figure out how to earn trust, rather than just claim it, will have a real advantage. Amodei seems to understand this. Whether the rest of the industry follows is the open question.

Sources

More from Future Technology