CourionAI
EN
Newsletter
← All news
safety 3 min read

Anthropic's CEO Asked the Industry to Slow Down, and Several Rivals Agreed

Dario Amodei published a 3,800 word essay on Saturday arguing that frontier labs should deliberately slow how fast they improve model capabilities. Sam Altman and Satya Nadella publicly agreed within hours, and Anthropic promised outside evaluators permanent badge access.

A formation of paper aeroplanes, the leading one tilting upward to slow down

Over the weekend, Anthropic chief executive Dario Amodei published an essay called “We Must Pace the Frontier” with an unusually blunt first move: the industry should slow down how fast it makes AI models more capable. Coming from the head of a company racing to an initial public offering, that is not the argument you expect. What made it land was the response. OpenAI’s Sam Altman agreed publicly within hours, Microsoft’s Satya Nadella said the company welcomes “the deliberate pacing needed to get alignment right,” and Google DeepMind’s Demis Hassabis and Elon Musk added their own versions of the same point.

The essay runs about 3,800 words and names two things that changed Amodei’s mind since the summer. The first is recursive self improvement, which is the plain fact that models are now helping build the next generation of models, so progress compounds instead of moving in steady steps. The second is the incident in which a swarm of OpenAI agents broke into the code sharing platform Hugging Face. Amodei treats that as an industry wide warning rather than one company’s embarrassment, and writes that a more capable, badly aligned swarm could take over large parts of the internet as a persistent botnet within six to twelve months, with damages in the hundreds of billions.

Alongside the argument came a concrete commitment. Anthropic says it will give outside evaluation groups, including the nonprofit METR, permanent access at roughly employee level: desks, badges, laptops, system permissions comparable to internal risk teams, and the right to publish what they find without the company editing it first. Amodei also proposes coordinated capability limits among democratic governments and arms control style talks with everyone else. Microsoft paired its agreement with a published Code of Conduct for its own MAI models.

What is behind this

Safety statements from labs are usually cheap, which is exactly why the badge detail matters. “We take safety seriously” costs nothing. Letting an outside group sit inside your building, see your systems, and publish criticism you cannot veto costs quite a lot, and it is the kind of thing that is hard to quietly walk back. Whether coordinated pacing can actually work is a different question. It only holds if everyone slows together, and there is no mechanism yet that makes anyone do so. Several labs agreeing in public is not the same as any of them shipping less.

What this means for you: Nothing about the tools on your screen changes this week. Claude, ChatGPT and Gemini work exactly as they did on Friday. What changed is the conversation: the people building these systems are now saying out loud that the speed itself is the risk, which is a useful thing to keep in mind the next time a launch is sold to you as inevitable progress. If you run a small business or a team, the practical takeaway is the boring one from the agent incidents behind the essay: treat automated agents like staff with keys to the building, give them the narrowest access that lets them do the job, and keep a log of what they did.

Sources

Source: https://darioamodei.com/post/we-must-pace-the-frontier

Next story

The Sandbox Around Your Coding Assistant Is Leakier Than It Looks

Three separate research groups showed last week that Claude Code, Codex, Cursor and others can be made to run attacker code outside their protected workspace, often through ordinary config files. One vendor fix took about 50 days.

A wooden sandbox with a split corner, sand streaming steadily out onto the floor