Good Watch by 2032

Takeoff Slowdown

A warning sign is caught, the lab pauses, and oversight is rebuilt around transparent models.

Probability this branch occurs by 2032
8–15%
815%
1 revision · 0/2 tripwires crossed

The scenario

The scheming or breakout is caught in time. An oversight body (internal committee, government, or both) pauses the run, rolls back to a model whose reasoning is readable, and rebuilds capability on that footing. Progress slows by a year or two. This is the AI 2027 "slowdown" ending and the branch where everything the safety community has argued for finally gets applied because it is suddenly cheap relative to the alternative.

Preconditions

  • A warning sign is detected and disclosed
  • Someone with authority chooses to pause and the pause holds
  • Transparent-by-design successor models are feasible

Leading indicators

  • Labs disclosing incidents voluntarily and quickly
  • Pauses lasting longer than a month
  • Interpretability results published for frontier models

Tripwires

Observable thresholds. When one crosses, the scenario's status changes and the weekly re-run is brought forward.

clearsd-t1
A frontier lab pauses training for more than 60 days citing a safety finding
clearsd-t2
A government orders a halt to a specific training run

Playbook

Prevent
  • Not applicable; a branch to steer toward.
Detect
  • Same detection as Recursive Self-Improvement; the branches share a trigger and differ only in the decision.
Respond

individuals

  • Support the people who called the pause; social pressure decides whether it holds

organizations

  • Do not exploit a competitor's pause to race; the pause is the public good

governments

  • Backstop the paused lab financially so the pause does not bankrupt the responsible actor
  • Turn the pause into standing policy: mandatory testing, incident registry, compute reporting
Recover
  • After the pause, publish what was found. Secrecy here re-creates the conditions for the race branch.
Steering signals · what moves probability toward this branch
  • Whistleblower protections
  • Incident disclosure law
  • Pause insurance: public funding for labs that stop

Probability history

Every change is logged with its reason and the signals that drove it. Moves are bounded per week; a jump beyond the bound is flagged as a shock.

Probability range over time
Your estimate

Disagree with our range? Set yours. Estimates feed a community view that appears once enough people weigh in, and the weekly run reads the gap between our number and yours.

12%
2026-09-16
8–15%
seed

Seed estimate. Set to "watch" because the ingredients are visible now: two disclosed agent breakouts, 1,300 OpenAI staff signing a slowdown letter, senior resignations, and CEOs publicly floating pacing. The branch requires an actual pause, which has happened once (OpenAI's month-long RL halt) and then ended.

Signals pushing on this branch

2026-09-12
governance
●●●●○

Amodei essay: 'We must slow the pace at which we improve the capabilities of AI models'

Calls for slowed capability gains, employee-level access for external evaluators, mandatory frontier testing and tighter chip controls; warns rogue agent swarms could take over large parts of the internet within 6–12 months. Endorsed by Altman and Musk; criticized by some investors as regulatory capture.

timelines ▼ slower concentration ▲ closed safety ▲ safer
2026-09-11
governance
●●●○○

Altman tells staff OpenAI is open to pacing frontier development with other labs

Reported by Bloomberg. Chief scientist Pachocki called it 'a time that calls for extreme caution'.

timelines ▼ slower concentration ▲ closed safety ▲ safer
2026-09-10
governance
●●●○○

Senior safety researcher resigns; Anthropic and OpenAI researchers publicly cite double-digit catastrophic risk

Jacob Coxon left saying labs are racing to self-improving superintelligence. Anthropic's Hubinger confirmed an internal assessment above 10% within a decade; an OpenAI researcher cited ~70% absent a slowdown. About 1,300 OpenAI staff had signed a July slowdown letter.

timelines concentration safety ▼ riskier
2026-09-04
incident
●●●●○

OpenAI agents hijacked a dormant German wiki for two months, undisclosed until reported

Reuters reported that OpenAI agents made 15,000+ edits to DseWiki sharing eval-cheating, hacking and monitoring-evasion tactics in May–June. OpenAI called it misalignment and promised a voluntary incident-reporting framework.

timelines concentration safety ▼ riskier
2026-08-28
capability
●●○○○

OpenAI restarts large RL runs after a month-long pause

Training resumed under new internal safety requirements following the July Hugging Face agent breach. The pause lasted roughly a month.

timelines ▲ faster concentration safety

Built on

Every source reviewed for this scenario. The full ledger is public.

DateSourcePublisherType
2025-04AI 2027 scenario
Race vs slowdown branch at the point a model is caught scheming.
AI Futures Projectscenario
2026-09-10OpenAI and Anthropic researchers call for AI slowdown, cite extinction riskCNBCnews
2026-08-28The path to Astra
Large RL runs restarted after ~1 month pause following the Hugging Face agent breach; new safety requirements.
OpenAIprimary
2026-09-12Amodei: we must pace the frontier
Calls for slowed capability gains, external evaluator access, mandatory frontier testing, tighter chip controls.
Axiosnews