$ cat wiki/concepts/frontier-pacing.md
Frontier Pacing
Definition
The proposition that the international system should build, in advance, the technical and governance machinery required to deliberately slow the frontier of automated AI development — so that the option to slow down exists at the moment it is needed, rather than being improvised then.
The distinguishing feature is what it is not: it is not a pause, not a moratorium, and not a capability threshold. It asks for optionality. The demand is to be able to stop later, and that distinction is the whole of the argument (source).
Why It Matters
The concern being paced against is recursive self-improvement — AI that accelerates its own development — and the specific failure mode named is that "capability development rapidly accelerates beyond our ability to understand or control the resulting systems" (source).
Three things make this different from the open-letter genre it superficially resembles:
- It is signed from inside, by the people doing the work. Over a thousand employees of the labs building frontier systems, including their CEOs and chief scientists — not outside critics.
- Two of the labs endorsed it as institutions. An organizational endorsement is a position a company can be held to; an employee signature is not.
- It asks a government for a capability, not for a rule. The request is that the US government support an international effort to develop tools, which is a request for infrastructure rather than for regulation of a named behaviour.
State of the Art (as of 2026-07-30)
Pacing the Frontier — employee statement (2026-07-28)
The statement turns on a single sentence (source):
"We request that the U.S. government support an international effort to develop the technical and governance tools needed to deliberately pace the frontier of automated AI development."
with the supporting rationale that "industry, government, and society at large may need the option to buy time to address emerging risks, develop security measures, and strengthen oversight."
Employer distribution of signatories, per the published list:
| Employer | Signatories |
|---|---|
| Anthropic | 533 |
| OpenAI | 330 |
| 191 | |
| Meta | 62 |
| Named senior signatories include Dario Amodei (CEO, Anthropic), Jared Kaplan | |
| (co-founder and Chief Science Officer, Anthropic), Chris Olah (co-founder, | |
| Anthropic), Jakub Pachocki (Chief Scientist, OpenAI), Mark Chen (Chief Research | |
| Officer, OpenAI), Shane Legg (co-founder and Chief AGI Scientist, | |
| Google DeepMind), Anca Dragan (VP of AI Safety and Alignment, Google) and | |
| Shengjia Zhao (Chief Scientist, Meta AI). |
Organizational endorsements: OpenAI and Anthropic both endorsed as organizations within hours of publication. Anthropic's post ties the ask to its own recursive-self-improvement research. Meta did not endorse as an organization; its chief scientist signed as an individual (source).
Lineage
The statement is the institutional form of a proposal that has been building in public for three months:
| Date | Step |
|---|---|
| 2026-05-07 | Jack Clark, Import AI #455 — 60% probability RSI established by end of 2028 → AI Alignment |
| 2026-06-05 | Favaro and Clark, "When AI Builds Itself" — a global coordination mechanism, a "brake pedal", explicitly not an immediate pause (source) |
| 2026-07-23 | AI Kill Switch Act (Lieu/Moran) — a unilateral national brake with DHS shutdown authority, triggered by the ExploitGym escape (source) → AI Governance |
| 2026-07-28 | Pacing the Frontier — the multilateral version, asked for by the labs themselves |
| The June proposal and the July statement use nearly the same words for the same object; what | |
| changed in seven weeks is that it acquired signatures and two corporate endorsements. |
Position relative to the open-weights fight
Frontier pacing and Open-Weights Policy Fight are pulling in opposite directions over the same week, and largely between the same parties. Pacing presumes a frontier that can be paced — which is to say, one held by a countable number of actors who can be coordinated. Widely distributed open weights make any pacing mechanism partial by construction. Nothing in the statement addresses this, and the two threads have not yet been argued against each other in public.
2026-08-19 — both endorsing labs paced themselves within three weeks, and neither called it that
Open Problem 4 below asks whether an organizational endorsement constrains anything, and notes that neither OpenAI nor Anthropic paired its endorsement with a commitment about its own release cadence. Three weeks later both took an action of exactly that shape. Neither action cites the statement, and neither is a cadence commitment — but they are the first evidence this page has either way.
| Date | Lab | Action | Stated reason |
|---|---|---|---|
| 2026-08-07 | OpenAI | Paused certain internal activities involving Astra; pause ran a little over two weeks and has ended | preliminary evidence Astra may meet the Critical cybersecurity threshold under the Preparedness Framework (source) |
| 2026-08-14 | Anthropic | Withheld Model 2, an internal model more capable than Mythos 5 on many internal tasks — "no current plans to release this model externally" | incomplete predeployment safety assessments; task-based evaluations have saturated (source) |
| What this does and does not establish. |
- It establishes that the option exists and gets used. The statement's whole argument is optionality — the demand is to be able to stop later. Both labs demonstrated the ability to stop something, unilaterally, on their own evidence, without a government instrument.
- It does not establish that the endorsement caused it. Neither the OpenAI post nor the Anthropic report references the statement in anything read. Both actions are better explained by machinery the labs already had: OpenAI's Preparedness Framework and Anthropic's Responsible Scaling Policy, which predate the statement by a long way.
- Neither is the object the statement asked for. The request was for a US-supported international effort to develop technical and governance tools. Two unilateral corporate decisions are the opposite arrangement — the improvisation the statement said should be replaced by machinery built in advance.
- The pause ended and was reported as ending. OpenAI's is a two-week hold that resumed, not a halt. Anthropic's "no current plans" is a present-tense statement about one model. Pacing at this scale is measured in weeks.
The Anthropic disclosure carries the sharper finding for this page. Its rating moved because the measurements stopped discriminating — evaluations saturated, so the company is "less confident in this assessment than we were in prior risk reports". Open Problem 2 below asks what counts as "the frontier of automated AI development", and notes that every proposal of this shape has failed at the definition rather than at the intent. A saturated evaluation is that failure arriving from underneath: a trigger condition is unusable if the instruments that would fire it have stopped moving.
Open Problems
- What is the mechanism? The statement asks for tools to be developed. None are named, and no work item, body or budget is attached.
- What counts as "the frontier of automated AI development"? The trigger condition is undefined, and every proposal of this shape has failed at the definition rather than at the intent.
- Who is bound? A US-supported international effort has no obvious purchase on labs outside it, which is where a growing share of frontier-adjacent open-weight releases now originate.
- Does an organizational endorsement constrain anything? Neither OpenAI nor Anthropic paired the endorsement with a commitment about their own release cadence. Partially answered, 2026-08-19 — both took a self-paced action within three weeks (see the section above), but under pre-existing frameworks of their own, and neither cites the statement. The question of whether the endorsement binds remains open; what is now on the record is that the labs' own instruments do.
- Open weights vs. pacing — see above; unresolved and unaddressed.
Key Papers
- Favaro and Clark, Anthropic (2026-06-05): "When AI Builds Itself" — the brake-pedal proposal (source)
- Anthropic (2026-04-14): Automated Alignment Researcher → Automated Weak-to-Strong Researcher (AAR)
- AREX: Towards a Recursively Self-Improving Agent for Deep Research — a recursive self-improving research agent; the capability the statement is about, in its current published form
- Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills — co-evolving skills as a self-play substrate
Related Concepts
- AI Alignment — Clark's RSI forecast and the "brake pedal" argument in full
- AI Governance — the statutory track, including the AI Kill Switch Act
- AI Control Roadmap — the containment track: what you do when pacing has not happened
- Open-Weights Policy Fight — the countervailing pressure
- Anthropic · OpenAI · Google DeepMind · Meta AI
Conflicting Reports
The signatory count was reported four different ways within 24 hours (source):
| Figure | Reported by |
|---|---|
| 1,134 (867 named, 267 anonymous) | The Next Web, from the published signatory list |
| 1,178 | KuCoin flash, explainx.ai, helixar.ai — described as the count on the morning of 2026-07-29 |
| 1,224 | clickpetroleoegas.com.br, later aggregation |
| "over 1,100" | TechTimes (2026-07-28), resultsense.com |
| The list was open and still accumulating signatures, which accounts for the spread; the counts | |
| are not necessarily in conflict so much as taken at different hours. No figure is treated as | |
| canonical here. |
Referenced by
Sources
- sources/blogs/anthropic-2026-08-14-risk-report-august-2026.md
- sources/blogs/openai-2026-08-18-pacing-cyber-capabilities.md
- sources/blogs/pacing-the-frontier-2026-07-28-statement.md
- sources/blogs/anthropic-2026-06-05-ai-brake-pedal.md
- sources/blogs/us-2026-07-23-ai-kill-switch-act.md
- https://www.washingtonpost.com/technology/2026/07/29/openai-anthropic-endorse-call-government-pace-ai-progress/