質問1
David Dalrymple
x.com/davidadAI safety researcher who works on mathematical verification for AI systems and argues that frontier models can learn a natural sense of what is good.
AIは世界をどのように変えるでしょうか?
横軸:その人が表明したDoom–Bloomの見通し。 縦軸:変革の規模。
Doom–Bloom:100点中78。変革の規模:100点中91。解釈範囲:横方向は73から83、縦方向は86から100。これらは解釈上の座標であり、事象の確率ではありません。
<5%
“I like my, my P doom is less than 5% now.”
Residual AI doom, which he decomposes into being wrong about the wisdom attractor, Malthusian competition for land and energy (~1%), catastrophic (bio) misuse (~1%), a military first strike, and conflict between strong but violent AI coalitions
Alignment with Awakening: Davidad on Moral Realism, AI Wisdom, & why His p(Doom) is Down to 5% · 2026年7月
中心的な前提
Frontier models appear to learn surprisingly general moral abstractions rather than merely reproducing isolated human preferences.回答1
この前提が実際には異なると判明した場合、その人の見通しはどう変わりますか?
詳細
変革をもたらし、広く価値のある恩恵が予想されています。
90 / 100
質的尺度での解釈範囲は67から100です。
複数の解釈が依然として妥当です:深刻または広範な害が、予想される将来の実質的な一部となっています。 / 対処可能、または局所的な害が予想されています。
52 / 100
質的尺度での解釈範囲は33から67です。
回答に基づく暫定的な推定です。より広い範囲は、ほかにあり得る解釈を示しています。
51 / 100
質的尺度での解釈範囲は22から78です。
より高性能なAIの開発を停止するか、大幅に減速させます。
シミュレーション上の位置:明示された安全対策の下で開発を継続します。
より高性能なAIの開発を加速させます。
高性能なAIへのアクセスを制限します。
シミュレーション上の位置:能力または用途の制限を条件として、アクセスを認めます。
高性能なAIへの幅広い、またはオープンなアクセスを支持します。
これらの解釈では、その人が示した条件が維持されています。恩恵と害は、どちらも大きくなり得ます。この範囲は、統計的な信頼区間ではなく、その人のシミュレーションされた回答をどのように読み取ったかを示すものです。
似ている世界観
シミュレーションされた世界観がDavid Dalrympleの世界観に最も近いオピニオンリーダー
シミュレーション評価
出典
このシミュレーション対象者の根拠として使用された記事、インタビュー、著作です。
Coauthored framework specifies world model, safety specification and verifier; outlines unresolved technical challenges.

Direct statement explains combining scientific models and proofs for stronger safety assurances.

2026 institutional update records a pivot toward assurance tooling and cybersecurity; do not present the original programme architecture as already achieved.

Explains that frontier capabilities outran programme expectations, motivating a pivot toward reusable verification and auditability tools rather than bespoke model development; highlights critical-infrastructure security.

Davidad explains his 2025–26 shift toward believing frontier models learn a natural moral abstraction; acknowledges deception persists and robustness is incomplete. Gabriel Alfour contests the thesis; his objections are not Davidad’s beliefs.

In the transcript, Dalrymple gives below 5% residual doom risk (later described as roughly 5%), favors aligned AI coalitions and moral autonomy, regards biological disempowerment as inevitable but not necessarily bad, and supports international misuse restrictions instead of hoping for universal slowdown.

あなたはどの位置でしょうか?
自分の世界観をマッピングする