Scientific Acceleration
AI agents run experiments and design molecules; discovery outpaces the institutions that approve and deploy it.
The scenario
Autonomous research agents propose, run and analyse experiments in biology, chemistry and materials. Bottlenecks move from ideas to wet labs, clinical trials and regulators. The main risk is not the science but the same dual-use capability that the bio-misuse branch tracks.
Preconditions
- Agents reliable on multi-week research tasks
- Automated lab capacity
- Regulators adapt trial and approval processes
Leading indicators
- AI-attributed publications and patents
- Automated lab investment
- Regulatory pilot programs for AI-designed therapeutics
Tripwires
Observable thresholds. When one crosses, the scenario's status changes and the weekly re-run is brought forward.
Playbook
- Not applicable; manage the dual-use edge via the bio-misuse playbook.
- Track publication and trial data.
individuals
- Scientists: become the person who frames questions and validates; that role grows
organizations
- Pharma and materials firms: automated labs and AI-native pipelines are the moat
governments
- Regulator capacity is the bottleneck; fund it
- Public funding for open scientific models so gains are not captured
- Not applicable.
- Regulatory modernization
- Open scientific datasets and models
- Automated lab infrastructure funding
Probability history
Every change is logged with its reason and the signals that drove it. Moves are bounded per week; a jump beyond the bound is flagged as a shock.
Disagree with our range? Set yours. Estimates feed a community view that appears once enough people weigh in, and the weekly run reads the gap between our number and yours.
Seed estimate, on watch: Terminal-Bench-Science jumped from 24.7% to 52.6% in one release; Amodei's "country of geniuses" framing has 90% by 2035.
Built on
Every source reviewed for this scenario. The full ledger is public.
| Date | Source | Publisher | Type |
|---|---|---|---|
| 2026-09-01 | Claude Fable 5.1 and Mythos 5.1 Mythos limited to vetted users; Terminal-Bench 4.0 55.8/60.9%; OSWorld 2.0 77.9%. | Anthropic | primary |
| 2026-02-13 | Dario Amodei on Dwarkesh Podcast 90% 'country of geniuses in a datacenter' by 2035; strong hunch 1–3 years. | Dwarkesh Podcast | statement |