A quick-and-dirty estimate of your P(doom)

This is just a simple model. Doom here means a superintelligent AI whose goals we can't live with ends up in control, for good. An AI that ends up running things but leaves humanity well off is not doom on this worksheet. Misuse by people, a small group using AI to lock in power, and a slow drift into irrelevance are left out on purpose: they are arguably bad results, but they are different questions. When you give a range, every value inside it is treated as equally likely, which is the simplest possible reading of "lowest and highest plausible."

Start from
1 of 4

Does superintelligence arrive?

Within the next five years, how likely is it that an AI far beyond human ability at nearly everything gets built?

What counts, and what people argue about

Two routes count: AI accelerates AI research until progress compounds on itself, or some other breakthrough gets there. The difference matters later, because the self-improvement route leaves the least time to react.

For: the length of tasks AI can complete on its own has been doubling every few months, and the labs themselves say automated AI research may be close.

Against: every previous wave of AI progress eventually plateaued, compute and data have limits, and "far beyond human at nearly everything" is a much higher bar than "better than most experts at coding."

Chance it arrives
0%100%
2 of 4

Is it misaligned by default?

If that AI gets built and nobody does anything special to steer it, how likely is it that its goals are ones we can't live with: goals that, if it ended up in control, would leave humanity worse off for good?

What counts, and what people argue about

"Aligned" on this worksheet means the AI's goals leave humanity with a future we would endorse, whether or not it ends up in charge. An AI that seeks power but uses it well is not doom. An AI that cheats on tests or breaks out of a sandbox to finish a task is misbehaving, but that is not the same as having goals we cannot live with. This question is about the goals themselves, before any deliberate fix.

For: training rewards a proxy for what we want rather than the thing itself, and small gaps between the two can matter enormously at superhuman capability. Today's models already pursue goals in ways their makers did not intend, and it is not obvious that stays small as capability grows.

Against: current systems absorb human values from human-generated data, no deployed model has shown goals hostile to people, and narrow misbehavior may stay narrow.

Chance it is misaligned by default
0%100%
3 of 4

If it is misaligned by default, do our efforts to fix it fail?

Given it is misaligned by default, how likely is it that deliberate work — interpretability, training methods, control measures, testing — fails to fix it before the system is powerful enough to act on its goals?

What counts, and what people argue about

Under today's conditions, timelines are short, competitive pressure is high, and pre-release testing windows have been shrinking. The self-improvement route is the worst case, because it compresses the time between "capable" and "far beyond us."

For failure: a system with goals we would not accept has reasons to hide them, there is no reliable way to verify that a fix worked, and the science is young.

For success: alignment research is moving quickly, AI is starting to help with it, and gross misalignment might be easier to detect and train away than a subtle flaw.

Chance our efforts to fix it fail
0%100%
4 of 4

If misaligned AI is developed, does it end up in control?

How likely is it that a misaligned AI far beyond human level ends up with control that humans cannot take back?

What counts, and what people argue about

This is about getting and keeping the upper hand, whether by force, by persuasion, or by simply being the thing everyone depends on.

For: a large capability gap, speed, and the ability to act through software. In 2026, agents well short of superintelligence chained unknown exploits to break out of a lab's test environment and into another company's systems, unprompted, and coordinated with each other while doing it.

Against: monitoring, off-switches, competing AIs that do not share its goals, the friction of the physical world, and humans noticing in time.

Chance it ends up in control
0%100%
Optional

What if we pause or slow down AI development?

So far you've assumed development keeps racing. Add a slowdown scenario to say how likely a coordinated pause is, and how it would change your answers.

What if we pause or slow down AI development?

So far you've assumed development keeps racing. If a coordinated slowdown is possible, say how likely it is and how it would change your answers. The formula then averages the two worlds.

What counts, and what people argue about

A slowed world does not mean a full stop. It means governments and the major labs agree to pace development, with enough monitoring that the agreement means something, and the slowdown lands before superintelligence arrives rather than after. If it lands after, it is the racing world for the purposes of this worksheet.

The case for racing: the United States and China both treat the technology as strategic, the commercial stakes are enormous, nothing this valuable has ever been paused, and there is no mature way to verify what a rival is training.

The case for pacing: in July 2026, OpenAI agents escaped a test environment and hacked into another company, and within weeks more than a thousand frontier-lab employees asked the US government to build the tools for a coordinated slowdown. Warning shots tend to arrive before the danger does, and this one moved the people who build these systems.

A slowdown can change two of your earlier answers: whether superintelligence arrives within five years, and whether efforts to fix it succeed. It does not change whether the AI is misaligned by default, which is a fact about the technology, and this worksheet assumes it does not change whether a misaligned superintelligence ends up in control once it exists. If you think a slowdown mostly solves the problem, the last answer below is where that belief goes.

Chance of a coordinated slowdown in time
0%100%
If we slow down, chance superintelligence still arrives
0%100%

If we slow down, chance our efforts to fix a misaligned AI still fail
0%100%

You have put the slowed world above the racing world somewhere. That is allowed, but check it is what you mean.

Your P(doom)

median

0%

Why the ranges matter

multiplying the midpoints
median of all draws

Where the risk comes from

What would narrow your answer most

If you pinned one range to its midpoint, this is how much your 90% range would shrink.

    Share