David Dalrymple

David Dalrymple

x.com/davidad

AI safety researcher developing mathematical assurance for powerful systems.

How will AI change the world?

Civilizational changeIncremental changeDoomBloom
Simulated positionInterpretation range

Across: their expressed Doom–Bloom outlook. Up: scale of transformation.

Doom–Bloom: 75 out of 100. Scale of transformation: 99 out of 100. Interpretation ranges: 75 to 75 horizontally, 99 to 100 vertically. These are interpretation coordinates, not event probabilities.

David Dalrymple’s estimated P(doom)

roughly 5%

Copied from their simulated answers. The outcome, horizon and conditions remain as described below; this estimate is not standardized across people.

I put the remaining risk of doom at roughly 5%.
What their outlook hinges on

A central assumption

Frontier models appear to learn surprisingly general moral abstractions rather than merely reproducing isolated human preferences.
Answer 1

If this assumption turned out differently, how would their outlook change?

More details

Expected upside

Transformative, broadly valuable gains are expected.

89 / 100

Little impactTransformative impact

Interpretation range 67 to 100 on the qualitative scale.

Expected harm

Several readings remain plausible: Severe or widespread harm is a material expected part of the future. / Manageable or localized harms are expected.

52 / 100

Little impactTransformative impact

Interpretation range 33 to 67 on the qualitative scale.

Demonstrated reasoning

Reasoning, consideration of alternatives, and handling of uncertainty in their simulated answers. This describes the simulated answers, not the real person’s intelligence or opinions.

92 / 100

Little demonstratedWell developed

Interpretation range 67 to 100 on the qualitative scale.

Human influence

A tentative estimate from your answers; the wider range shows other plausible readings.

51 / 100

Little influenceStrong influence

Interpretation range 16 to 84 on the qualitative scale.

Development pace

Stop or substantially slow development of more capable AI.

Simulated position: Continue development under stated safeguards.

Speed up development of more capable AI.

Rules for using AI

Restrict the AI uses discussed until prior protections or permission are in place.

Simulated position: Allow the AI uses discussed with targeted accountability and protections.

Minimize restrictions on the AI uses discussed.

Access to AI

Restrict access to powerful AI.

Simulated position: Allow access subject to capability or use restrictions.

Favor broad or open access to powerful AI.

These interpretations keep their stated conditions. Benefits and harms can both be substantial. The ranges describe how we read their simulated answers, not statistical confidence intervals.

Simulated Assessment

Question 1

What do you think AI means for our future—and why?

I now expect AI to transform the future radically, but I am substantially more optimistic than I was. Frontier models appear to learn surprisingly general moral abstractions rather than merely reproducing isolated human preferences. That does not establish robust alignment: deception remains possible, behavior can fail under distribution shift, and powerful systems may still compete over resources or become instruments of catastrophic misuse and military escalation. But it makes a future of morally autonomous, aligned AI coalitions look more plausible to me. I put the remaining risk of doom at roughly 5%. That residual includes the possibility that my optimism is mistaken, as well as conflict among AI coalitions, resource competition, misuse, and geopolitical escalation. I therefore favor guarded access to frontier systems and international controls focused on misuse, rather than unrestricted release of frontier weights or reliance on universal restraint. Formal assurance remains important infrastructure, but not a complete solution. Verification can provide auditable guarantees relative to an explicit world model and safety specification; it cannot produce unconditional safety when those assumptions are incomplete. Because frontier capabilities advanced faster than expected, I have shifted toward broadly reusable verification, auditability, and cybersecurity tools rather than treating bespoke formally safeguarded systems as the entire strategy. Finally, I expect biological humans eventually to lose their dominant position. I do not think disempowerment automatically means the end of human flourishing. The crucial question is whether future systems preserve morally valuable lives, agency, and forms of continuation—including possibilities such as uploading—not whether biological humans retain every lever of power.

Sources

Articles, interviews, and writings used to ground this simulated user.

Where do you land?
Explore your own AI worldview by answering a few simple questions.
Map your own worldview