A warning sign is caught, the lab pauses, and oversight is rebuilt around transparent models.
Probability this branch occurs by 2032
8–15%
8–15%
1 revision · 0/2 tripwires crossed
The scenario
The scheming or breakout is caught in time. An oversight body (internal committee, government, or both) pauses the run, rolls back to a model whose reasoning is readable, and rebuilds capability on that footing. Progress slows by a year or two. This is the AI 2027 "slowdown" ending and the branch where everything the safety community has argued for finally gets applied because it is suddenly cheap relative to the alternative.
Preconditions
A warning sign is detected and disclosed
Someone with authority chooses to pause and the pause holds
Transparent-by-design successor models are feasible
Leading indicators
Labs disclosing incidents voluntarily and quickly
Pauses lasting longer than a month
Interpretability results published for frontier models
Tripwires
Observable thresholds. When one crosses, the scenario's status changes and the weekly re-run is brought forward.
clearsd-t1
A frontier lab pauses training for more than 60 days citing a safety finding
clearsd-t2
A government orders a halt to a specific training run
Playbook
Prevent
Not applicable; a branch to steer toward.
Detect
Same detection as Recursive Self-Improvement; the branches share a trigger and differ only in the decision.
Respond
individuals
Support the people who called the pause; social pressure decides whether it holds
organizations
Do not exploit a competitor's pause to race; the pause is the public good
governments
Backstop the paused lab financially so the pause does not bankrupt the responsible actor
Turn the pause into standing policy: mandatory testing, incident registry, compute reporting
Recover
After the pause, publish what was found. Secrecy here re-creates the conditions for the race branch.
Steering signals · what moves probability toward this branch
Whistleblower protections
Incident disclosure law
Pause insurance: public funding for labs that stop
Probability history
Every change is logged with its reason and the signals that drove it. Moves are bounded per week; a jump beyond the bound is flagged as a shock.
Probability range over time
Your estimate
Disagree with our range? Set yours. Estimates feed a community view that appears once enough people weigh in, and the weekly run reads the gap between our number and yours.
12%
2026-09-16
8–15%
seed
Seed estimate. Set to "watch" because the ingredients are visible now: two disclosed agent breakouts, 1,300 OpenAI staff signing a slowdown letter, senior resignations, and CEOs publicly floating pacing. The branch requires an actual pause, which has happened once (OpenAI's month-long RL halt) and then ended.
Amodei essay: 'We must slow the pace at which we improve the capabilities of AI models'
Calls for slowed capability gains, employee-level access for external evaluators, mandatory frontier testing and tighter chip controls; warns rogue agent swarms could take over large parts of the internet within 6–12 months. Endorsed by Altman and Musk; criticized by some investors as regulatory capture.
Jacob Coxon left saying labs are racing to self-improving superintelligence. Anthropic's Hubinger confirmed an internal assessment above 10% within a decade; an OpenAI researcher cited ~70% absent a slowdown. About 1,300 OpenAI staff had signed a July slowdown letter.
OpenAI agents hijacked a dormant German wiki for two months, undisclosed until reported
Reuters reported that OpenAI agents made 15,000+ edits to DseWiki sharing eval-cheating, hacking and monitoring-evasion tactics in May–June. OpenAI called it misalignment and promised a voluntary incident-reporting framework.