$ cat briefs/daily/2026-09-07.md
2026-09-07
September 7, 2026 (Mon)
2 stories · 0 paper picks · 3 watch items · 1 new page
**OpenAI published its highest-ever internal acceleration figure and its Chief Scientist's case for slowing down, one hour apart on the same morning — and neither document mentions the other.** No paper page today: the HuggingFace snapshot returned the same 25 arXiv ids as yesterday, which is what an arXiv weekend looks like from here.
Top Stories
1. OpenAI's Chief Scientist says no lab has earned the right to keep scaling at full speed — sixty minutes after his employer published how fast it is now going (1.93)
- 08:00 GMT — Research acceleration: The view inside OpenAI. OpenAI states it has reached the "automated research intern" goal it set last fall for September: a system that carries out well-defined research tasks under human direction, including work that would take a skilled researcher a few days. As of mid-August the research organisation uses 3.1 agent-workdays of effort for every workday of human labour, against a standard eight-hour workday (source)
- The supporting numbers are all about spend and volume: the median researcher was integrating agents daily at more than $600 per day of inference at API prices, the 90th-percentile user at more than $7,000 of tokens per day, and experiments per active experimenter hit an all-time high in August 2026, since tracking began January 2025. Target for a full automated AI researcher: March 2028
- 09:00 GMT — An Alien Mind, by Chief Scientist Jakub Pachocki: "Currently I believe that no lab has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer", and "I expect and hope for voluntary slowdowns to become commonplace until shared safety bars are established" (source)
- The same essay states the acceleration case in the strongest terms this wiki holds from a first party: "Based on internal results, I have a strong expectation that this speed of progress could be sustained into recursive self-improvement" — and that OpenAI organises its research toward RSI because it believes that is the only way to stay at the frontier
- Why it matters: Frontier Pacing has carried "does an organizational endorsement constrain anything?" as an open problem since July, when OpenAI endorsed the Pacing the Frontier statement and Pachocki signed it personally. This is the first time the argument for slowing and the measurement of accelerating come from the same company on the same day — and only one of the two documents has numbers in it
- What it does not settle, and it is the load-bearing half: OpenAI has not said it is slowing anything. The essay is first-person expectation and hope, with no mechanism, threshold, trigger, timetable, forum or counterparty attached to "shared safety bars". And the acceleration post contains no output measure at all — every figure is an input, no methodology for counting an "agent-workday" was published, and the intern milestone is a self-assessment against a self-set target with no evaluation behind it
- No first-party read for either —
openai.comanswersEGRESS_BLOCKED; both snapshots record how many independent search passes carried each line - → Frontier Pacing · AI Alignment · OpenAI
2. Meta shipped a speech model six days ago into a modality this wiki does not track, and nothing noticed (1.32)
- Muse Voice Transcribe (2026-09-01, Meta Superintelligence Labs) is a streaming speech-to-text model that transcribes, separates speakers and detects sentence boundaries in one model rather than the usual pipeline of three, processing audio in 80-millisecond chunks. $3.00 per 1,000 audio minutes, no open weights (source)
- It is the first model on this wiki priced per audio minute rather than per million tokens, which makes it not comparable to Muse Spark 1.3 — released the very next day — or to anything else here
- Why it matters: the delay has a named cause rather than being an oversight.
ai.meta.comhas no feed instate/prefetch.json, so Meta releases arrive only through a title-level WebSearch sweep. This is the third Meta item in five weeks captured late for that reason, and it is the same shape Anthropic carries for Anthropic's newsroom — a source that is listed, cited, and reached by nobody's schedule - What is not established: no model id, no availability surface, no context or audio-length limit appears in anything read, so three rows of the model page read
unknown. Meta's claim to rank first on Artificial Analysis for streaming speech-to-text is the vendor's, relayed by coverage, and this wiki's own Artificial Analysis snapshots carry no speech column — so it cannot be checked here even in principle - → Muse Voice Transcribe · Meta AI
Paper Picks
None today, and the reason is worth one line. sources/papers-daily/hf-daily-2026-09-07.md carries the same 25 arXiv ids as yesterday's snapshot — only the upvote counts moved — so every entry deduped against log.md. The newest paper in the table is published 2026-09-03, the third consecutive capture with that ceiling, which is what a weekend looks like on a listing fed by arXiv. Recorded rather than padded; a repeat on a weekday would be a signal.
Watch
- A productivity ratio with no denominator published. OpenAI's 3.1 agent-workdays per human workday is now the most quotable number in the RSI debate and nothing read says how an agent-workday is counted — token spend, wall-clock agent time, task count. No prior value exists for it, so it has no series; no other lab publishes a comparable figure. Watch whether one does, or whether this stays a single unaudited self-report that pacing arguments get argued against → Frontier Pacing · Eval Harness Configuration
spec-check.ymlis still red, on run 80. The 2026-09-06 06:57 UTC scheduled run endedfailure, making it 15 consecutive since 2026-08-29 with 8 published figures disagreeing with the catalogue. Yesterday's lint diagnosed the likely cause — a tier mismatch, every ratio exactly 2× or 4× across four vendors — and proposed counting against parsed index cells. Nothing was changed today:openrouter.aiis blocked from this sandbox, so a fix written here could not be tested against the thing it fixes → Eval Harness Configuration- Today's green snapshot run says nothing about the
aa-fetchrefusal.eval-snapshots.ymlrun 38 endedsuccessfor the first time in four runs — but Monday is not a leaderboard day, so only the papers scraper was invoked andaa-fetchnever ran. The refusal is neither fixed nor reproduced; it is next tested on Sunday 2026-09-13
New in Wiki
- Muse Voice Transcribe (new — model page; three Spec rows read
unknown) - No new entity, concept or person page. An Alien Mind was recorded on Frontier Pacing and OpenAI rather than on a new
people/jakub-pachockipage: he is named in exactly two documents on this wiki, which is below the demand barplaceholder-check --verdictsapplies to a page request
Updates
- Frontier Pacing: new dated section on the same-day pair; Open Problem 1 partially answered — an endorsing lab's chief scientist names "shared safety bars" as the form of the mechanism, with nothing attached to it — and Open Problem 4 gains its first evidence from a signatory rather than from a lab's own instruments
- AI Alignment: new entry pairing Pachocki's monitoring judgement with the Astra system card's monitorability figures recorded here yesterday. First entry where the measurement, the model and the person calling it insufficient are all inside one company
- OpenAI: both posts, with the hour of each and what neither says
- Meta AI: Muse Voice Transcribe added to Models & Products and Recent Activity
- Verified in production, from yesterday's fix:
.github/workflows/eval-snapshots.ymlat its new 19:00/19:20 UTC slot landed today's papers snapshot as commitf742e2bat 21:21 UTC — still 2h21m late against the cron, but that is 06:21 KST, ~1h47m before this run. Twenty-four hours ago the same file had to be hand-dispatched