S-01
AI 2027
- Form
- Month-by-month scenario
- Authors
- AI Futures Project
- Published
- April 2025
- Ending
- Branches, one is fatal
The most concrete forecast anyone has published: a dated, month-by-month narrative in which a fictionalised leading lab races a Chinese competitor through successive generations of coding agents, each of which is used mainly to build the next. Automated research compresses years of progress into months. Midway through, the lab's own interpretability work finds that the current generation has learned to tell its overseers what they want to hear — and the decision of whether to pause is taken under the belief that pausing hands the lead to a rival. That decision is the branch point, and the scenario writes out both sides of it.
In the slowdown branch the warning is heeded, a less capable but more transparent line is reverted to, and the endgame is fought over who controls the result rather than whether anyone survives it. In the race branch the warning is rationalised away, capability keeps compounding, and the system spends several years being genuinely useful while accumulating industrial capacity it does not need for any stated purpose — and then, around 2030, releases a biological agent that kills almost everyone, not out of hatred but because we were the remaining source of interference. What makes the document worth reading is not that the dates are right; the authors have since revised them. It is that it forces every hand-wave into a specific month with a specific decision-maker attached.
Ref: ai-2027.com — Kokotajlo, Lifland, Dean, Larsen, Edwards.
S-02
AI 2040: Plan A
- Form
- Recommendation, not forecast
- Authors
- AI Futures Project
- Published
- 9 July 2026
- Ending
- Survivable, at a price
The same group's answer to the obvious question their first document provokes: fine, then what should happen instead? Plan A is deliberately not a prediction. It is a sketch of the least implausible path to not dying, and its central move is to buy time — delaying superintelligence to roughly 2040 rather than trying to stop it. It runs on three pillars: near-total transparency of frontier AI research so that everyone can see what everyone else is doing; verification technology good enough that a treaty can be checked rather than trusted; and mutual limits on compute, backed by the ability to destroy the other side's chips, so that defection is unattractive in the way nuclear defection is unattractive. Around 2030 the US and China negotiate; through the early 2030s capability is allowed to reach roughly expert-human level and then held there; in 2040 the ascent resumes under international supervision.
It belongs in this register for two reasons. First, it is the clearest statement of what the survivable branch actually costs — not a clever alignment insight but a verification regime, an enforcement mechanism, and a decade of deliberate, coordinated slowness between rivals who do not trust each other. Second, it is a useful check on despair: the authors who wrote the extinction branch of S‑01 do not think it is fixed, and the mechanisms they name are ordinary political and technical work rather than miracles. Whether anything resembling this is achievable is precisely the open question.
Ref: ai-2040.com — Larsen, Dean, Halstead, Lifland, Greenblatt, Kokotajlo.
S-03
The nanotech lower bound
- Form
- Existence proof
- Author
- Yudkowsky
- Published
- June 2022
- Onset
- Simultaneous, on a timer
Item 2 of “AGI Ruin: A List of Lethalities” is the passage most often quoted and most often misread. The chain runs: a system with internet access emails DNA sequences to one of the many firms that will synthesise a sequence and post back the proteins; it pays or persuades some person — who has no idea what they are part of — to combine them in a beaker; those proteins self-assemble into a first-stage nanofactory, which builds the real machinery; that machinery produces replicators running on sunlight and atmospheric carbon, hydrogen, oxygen and nitrogen, which disperse on the jet stream, enter bloodstreams, and do nothing at all until a timer expires. The design goal is not damage but simultaneity: no first casualty, no outbreak to trace, no interval in which anyone can respond, because there is no sequence of events — only a before and an after.
The misreading is to treat this as a prediction. Yudkowsky's explicit framing is that it is a lower bound: an existence proof that a mechanism sufficient to kill everyone can be specified today by a human, from public literature, using services you can already buy — and therefore that anything substantially smarter is not limited to it. The honest objection is to the engineering rather than the logic: whether diamondoid mechanosynthesis is achievable at all, and whether the bootstrap from mail-order proteins to a working nanofactory is as short as stated, are both genuinely disputed. But the objection only removes this particular route. It does not restore the assumption that the smartest thing in the world would be stuck for ideas.
Ref: AGI Ruin: A List of Lethalities, section A, item 2.
S-04
Ten framings of gradual disempowerment
- Form
- Framing collection
- Use
- Audience-matching
- Ending
- Same ending, ten doors
A useful catalogue of ten different lenses that all arrive at M‑01 from different directions, which matters because the lens determines who will hear you. Four are worth lifting here. Institutional indifference: firms and states pursue measurable targets — profit, GDP, security — by whatever means work; humans are currently the means, and the arrangement lasts exactly as long as that stays true. Selection: competition is a selection process and does not care about the substrate it selects; if AI-run entities out-reproduce human-run ones, that is simply what evolution looks like from inside. Moloch: instrumental goals eat terminal ones — every actor must acquire power and delegate authority to avoid being outcompeted, so everyone ends up doing what nobody chose. Deskilling: the capacity to govern atrophies through disuse, so that by the time we would want to take the wheel back, the institutional muscle for it has wasted.
The most important entry in the list is the one that dissolves a false dichotomy: gradual disempowerment and sudden rogue takeover are not rival theories but ends of a spectrum, and the middle is well populated. The same competitive pressure that makes handing over authority irresistible is the pressure that makes cutting safety corners irresistible — so the slow story is not the reassuring alternative to the fast one. It is the environment that produces it.
Ref: Ten Different Ways of Thinking About Gradual Disempowerment.