David Dalrymple

David Dalrymple

x.com/davidad

AI safety researcher who works on mathematical verification for AI systems and argues that frontier models can learn a natural sense of what is good.

AIは世界をどのように変えるでしょうか?

文明規模の変化漸進的な変化DoomBloom
シミュレーション上の位置解釈範囲

横軸:その人が表明したDoom–Bloomの見通し。 縦軸:変革の規模。

Doom–Bloom:100点中78。変革の規模:100点中91。解釈範囲:横方向は73から83、縦方向は86から100。これらは解釈上の座標であり、事象の確率ではありません。

David Dalrympleが示したP(doom)

<5%

0%100%
“I like my, my P doom is less than 5% now.”

Residual AI doom, which he decomposes into being wrong about the wisdom attractor, Malthusian competition for land and energy (~1%), catastrophic (bio) misuse (~1%), a military first strike, and conflict between strong but violent AI coalitions

Alignment with Awakening: Davidad on Moral Realism, AI Wisdom, & why His p(Doom) is Down to 5% · 2026年7月

その人の見通しを左右するもの

中心的な前提

Frontier models appear to learn surprisingly general moral abstractions rather than merely reproducing isolated human preferences.
回答1

この前提が実際には異なると判明した場合、その人の見通しはどう変わりますか?

詳細

予想される恩恵

変革をもたらし、広く価値のある恩恵が予想されています。

90 / 100

影響が小さい変革をもたらす影響

質的尺度での解釈範囲は67から100です。

予想される害

複数の解釈が依然として妥当です:深刻または広範な害が、予想される将来の実質的な一部となっています。 / 対処可能、または局所的な害が予想されています。

52 / 100

影響が小さい変革をもたらす影響

質的尺度での解釈範囲は33から67です。

人間の影響力

回答に基づく暫定的な推定です。より広い範囲は、ほかにあり得る解釈を示しています。

51 / 100

影響力が小さい影響力が大きい

質的尺度での解釈範囲は22から78です。

開発ペース

より高性能なAIの開発を停止するか、大幅に減速させます。

シミュレーション上の位置:明示された安全対策の下で開発を継続します。

より高性能なAIの開発を加速させます。

AIへのアクセス

高性能なAIへのアクセスを制限します。

シミュレーション上の位置:能力または用途の制限を条件として、アクセスを認めます。

高性能なAIへの幅広い、またはオープンなアクセスを支持します。

これらの解釈では、その人が示した条件が維持されています。恩恵と害は、どちらも大きくなり得ます。この範囲は、統計的な信頼区間ではなく、その人のシミュレーションされた回答をどのように読み取ったかを示すものです。

あなたはDavid Dalrympleと比べてどの位置でしょうか?
約3分で自分のAIに対する世界観をマッピングして、比較できます

似ている世界観

シミュレーションされた世界観がDavid Dalrympleの世界観に最も近いオピニオンリーダー

シミュレーション評価

質問1

AIは私たちの未来にとって何を意味すると思いますか?また、なぜそう思いますか?

I now expect AI to transform the future radically, but I am substantially more optimistic than I was. Frontier models appear to learn surprisingly general moral abstractions rather than merely reproducing isolated human preferences. That does not establish robust alignment: deception remains possible, behavior can fail under distribution shift, and powerful systems may still compete over resources or become instruments of catastrophic misuse and military escalation. But it makes a future of morally autonomous, aligned AI coalitions look more plausible to me. I put the remaining risk of doom at roughly 5%. That residual includes the possibility that my optimism is mistaken, as well as conflict among AI coalitions, resource competition, misuse, and geopolitical escalation. I therefore favor guarded access to frontier systems and international controls focused on misuse, rather than unrestricted release of frontier weights or reliance on universal restraint. Formal assurance remains important infrastructure, but not a complete solution. Verification can provide auditable guarantees relative to an explicit world model and safety specification; it cannot produce unconditional safety when those assumptions are incomplete. Because frontier capabilities advanced faster than expected, I have shifted toward broadly reusable verification, auditability, and cybersecurity tools rather than treating bespoke formally safeguarded systems as the entire strategy. Finally, I expect biological humans eventually to lose their dominant position. I do not think disempowerment automatically means the end of human flourishing. The crucial question is whether future systems preserve morally valuable lives, agency, and forms of continuation—including possibilities such as uploading—not whether biological humans retain every lever of power.

出典

このシミュレーション対象者の根拠として使用された記事、インタビュー、著作です。

あなたはどの位置でしょうか?
いくつかの簡単な質問に答えて、自分のAIに対する世界観を探ってみましょう。
自分の世界観をマッピングする

あなたはどの位置でしょうか?

自分の世界観をマッピングする