AI Trend Notifier
EN
← wiki

$ cat wiki/concepts/frontier-pacing.md

Frontier Pacing

conceptupdated 2026-08-19created 2026-07-30

Definition

The proposition that the international system should build, in advance, the technical and governance machinery required to deliberately slow the frontier of automated AI development — so that the option to slow down exists at the moment it is needed, rather than being improvised then.

The distinguishing feature is what it is not: it is not a pause, not a moratorium, and not a capability threshold. It asks for optionality. The demand is to be able to stop later, and that distinction is the whole of the argument (source).

Why It Matters

The concern being paced against is recursive self-improvement — AI that accelerates its own development — and the specific failure mode named is that "capability development rapidly accelerates beyond our ability to understand or control the resulting systems" (source).

Three things make this different from the open-letter genre it superficially resembles:

  1. It is signed from inside, by the people doing the work. Over a thousand employees of the labs building frontier systems, including their CEOs and chief scientists — not outside critics.
  2. Two of the labs endorsed it as institutions. An organizational endorsement is a position a company can be held to; an employee signature is not.
  3. It asks a government for a capability, not for a rule. The request is that the US government support an international effort to develop tools, which is a request for infrastructure rather than for regulation of a named behaviour.

State of the Art (as of 2026-07-30)

Pacing the Frontier — employee statement (2026-07-28)

The statement turns on a single sentence (source):

"We request that the U.S. government support an international effort to develop the technical and governance tools needed to deliberately pace the frontier of automated AI development."

with the supporting rationale that "industry, government, and society at large may need the option to buy time to address emerging risks, develop security measures, and strengthen oversight."

Employer distribution of signatories, per the published list:

EmployerSignatories
Anthropic533
OpenAI330
Google191
Meta62
Named senior signatories include Dario Amodei (CEO, Anthropic), Jared Kaplan
(co-founder and Chief Science Officer, Anthropic), Chris Olah (co-founder,
Anthropic), Jakub Pachocki (Chief Scientist, OpenAI), Mark Chen (Chief Research
Officer, OpenAI), Shane Legg (co-founder and Chief AGI Scientist,
Google DeepMind), Anca Dragan (VP of AI Safety and Alignment, Google) and
Shengjia Zhao (Chief Scientist, Meta AI).

Organizational endorsements: OpenAI and Anthropic both endorsed as organizations within hours of publication. Anthropic's post ties the ask to its own recursive-self-improvement research. Meta did not endorse as an organization; its chief scientist signed as an individual (source).

Lineage

The statement is the institutional form of a proposal that has been building in public for three months:

DateStep
2026-05-07Jack Clark, Import AI #455 — 60% probability RSI established by end of 2028 → AI Alignment
2026-06-05Favaro and Clark, "When AI Builds Itself" — a global coordination mechanism, a "brake pedal", explicitly not an immediate pause (source)
2026-07-23AI Kill Switch Act (Lieu/Moran) — a unilateral national brake with DHS shutdown authority, triggered by the ExploitGym escape (source) → AI Governance
2026-07-28Pacing the Frontier — the multilateral version, asked for by the labs themselves
The June proposal and the July statement use nearly the same words for the same object; what
changed in seven weeks is that it acquired signatures and two corporate endorsements.

Position relative to the open-weights fight

Frontier pacing and Open-Weights Policy Fight are pulling in opposite directions over the same week, and largely between the same parties. Pacing presumes a frontier that can be paced — which is to say, one held by a countable number of actors who can be coordinated. Widely distributed open weights make any pacing mechanism partial by construction. Nothing in the statement addresses this, and the two threads have not yet been argued against each other in public.

2026-08-19 — both endorsing labs paced themselves within three weeks, and neither called it that

Open Problem 4 below asks whether an organizational endorsement constrains anything, and notes that neither OpenAI nor Anthropic paired its endorsement with a commitment about its own release cadence. Three weeks later both took an action of exactly that shape. Neither action cites the statement, and neither is a cadence commitment — but they are the first evidence this page has either way.

DateLabActionStated reason
2026-08-07OpenAIPaused certain internal activities involving Astra; pause ran a little over two weeks and has endedpreliminary evidence Astra may meet the Critical cybersecurity threshold under the Preparedness Framework (source)
2026-08-14AnthropicWithheld Model 2, an internal model more capable than Mythos 5 on many internal tasks — "no current plans to release this model externally"incomplete predeployment safety assessments; task-based evaluations have saturated (source)
What this does and does not establish.
  • It establishes that the option exists and gets used. The statement's whole argument is optionality — the demand is to be able to stop later. Both labs demonstrated the ability to stop something, unilaterally, on their own evidence, without a government instrument.
  • It does not establish that the endorsement caused it. Neither the OpenAI post nor the Anthropic report references the statement in anything read. Both actions are better explained by machinery the labs already had: OpenAI's Preparedness Framework and Anthropic's Responsible Scaling Policy, which predate the statement by a long way.
  • Neither is the object the statement asked for. The request was for a US-supported international effort to develop technical and governance tools. Two unilateral corporate decisions are the opposite arrangement — the improvisation the statement said should be replaced by machinery built in advance.
  • The pause ended and was reported as ending. OpenAI's is a two-week hold that resumed, not a halt. Anthropic's "no current plans" is a present-tense statement about one model. Pacing at this scale is measured in weeks.

The Anthropic disclosure carries the sharper finding for this page. Its rating moved because the measurements stopped discriminating — evaluations saturated, so the company is "less confident in this assessment than we were in prior risk reports". Open Problem 2 below asks what counts as "the frontier of automated AI development", and notes that every proposal of this shape has failed at the definition rather than at the intent. A saturated evaluation is that failure arriving from underneath: a trigger condition is unusable if the instruments that would fire it have stopped moving.

Open Problems

  1. What is the mechanism? The statement asks for tools to be developed. None are named, and no work item, body or budget is attached.
  2. What counts as "the frontier of automated AI development"? The trigger condition is undefined, and every proposal of this shape has failed at the definition rather than at the intent.
  3. Who is bound? A US-supported international effort has no obvious purchase on labs outside it, which is where a growing share of frontier-adjacent open-weight releases now originate.
  4. Does an organizational endorsement constrain anything? Neither OpenAI nor Anthropic paired the endorsement with a commitment about their own release cadence. Partially answered, 2026-08-19 — both took a self-paced action within three weeks (see the section above), but under pre-existing frameworks of their own, and neither cites the statement. The question of whether the endorsement binds remains open; what is now on the record is that the labs' own instruments do.
  5. Open weights vs. pacing — see above; unresolved and unaddressed.

Key Papers

Conflicting Reports

The signatory count was reported four different ways within 24 hours (source):

FigureReported by
1,134 (867 named, 267 anonymous)The Next Web, from the published signatory list
1,178KuCoin flash, explainx.ai, helixar.ai — described as the count on the morning of 2026-07-29
1,224clickpetroleoegas.com.br, later aggregation
"over 1,100"TechTimes (2026-07-28), resultsense.com
The list was open and still accumulating signatures, which accounts for the spread; the counts
are not necessarily in conflict so much as taken at different hours. No figure is treated as
canonical here.

Referenced by

Sources