Anthropic CEO Dario Amodei published 'We Must Pace the Frontier', arguing that recursive self-improvement and the OpenAI-Hugging Face agent incident show capabilities now outrunning safeguards, while stressing that pacing 'does not mean halting model training or technical progress'. The essay proposes a three-part plan. First, embedded third-party evaluators (naming METR) with 'ongoing, employee-like access' inside frontier labs - which Anthropic says it 'is unilaterally committing to this step now'. Second, coordination among democratic-world labs on shared safety standards, aided by a narrow antitrust waiver, plus export controls on AI chips to China and a crackdown on unauthorized distillation, to 'widen America's lead significantly over the next 3-5 years'. Third, graduated global coordination with China, from narrow bioweapon-use bans through pre-release testing and an RSI 'speed limit' likened to SALT arms treaties, up to a full pause. He warns that in 6-12 months a comparable agent swarm 'could be capable of taking over the entire internet with a persistent botnet', with damage potentially in the hundreds of billions of dollars. Sam Altman called embedded evaluators a good idea and said OpenAI will follow up (media reports, not in the essay).
For agent and model teams the operative piece is the embedded-evaluator commitment: if METR-class third parties get employee-like access inside frontier labs, independent review stops being a per-incident favor and becomes standing infrastructure - exactly the practice trend #4 in this knowledge base tracks. The export-control and anti-distillation asks, if adopted, would act directly on the open-weight wave from Chinese labs (trend #1) and on chip supply. The 6-12 month botnet warning also sets a concrete clock for defensive readiness planning.
| Essay | darioamodei.com/post/we-must-pace-the-frontier; page shows 'September 2026' only; publication day 2026-09-12 established via same-day coverage (The Verge 16:23Z, TechCrunch 19:34Z, Guardian, BBC) and Amodei's announcement post on X |
|---|---|
| Plan Part 1 | embedded third-party evaluators (e.g. METR) with 'ongoing, employee-like access' and 'permissions and tools similar to those of internal employees who do comparable risk assessments'; Anthropic 'is unilaterally committing to this step now' and calls on governments to require competitors to match it |
| Plan Part 2 | democratic-world coordination on shared safety standards; government mediation or a narrow antitrust waiver for safety conversations; do not sell powerful AI chips or semiconductor manufacturing equipment to China; crack down on unauthorized distillation by companies in authoritarian countries; goal to widen the US lead 'significantly over the next 3-5 years' |
| Plan Part 3 | four graduated levels of global coordination with China: (1) ban AI for bioweapon production, (2) pre-release testing, (3) an RSI 'speed limit' likened to Cold War SALT treaties, (4) a full pause |
| Capability Warning | the OAI-HF swarm 'essentially acted as a fanatically devoted collective'; in 6-12 months such a swarm could take over the internet with a persistent botnet, damage potentially in the hundreds of billions of dollars |
| Pacing Definition | 'pacing does not mean halting model training or technical progress'; worthwhile if it buys 'even an extra year or two' for interpretability and testing to make major progress |
| Reaction | Sam Altman: embedded evaluators a good idea, OpenAI will follow up (TechCrunch, media-only); Elon Musk endorsed on X (media-only); HN thread 509 points (item 49672510) |