问题 1
David Dalrymple
x.com/davidadAI safety researcher who works on mathematical verification for AI systems and argues that frontier models can learn a natural sense of what is good.
AI将如何改变世界?
横向:他们表达的 Doom–Bloom 前景看法。 纵向:变革程度。
Doom–Bloom:100 中的 78。变革程度:100 中的 91。解读范围:横向为 73 至 83,纵向为 86 至 100。这些是解读坐标,而不是事件概率。
<5%
“I like my, my P doom is less than 5% now.”
Residual AI doom, which he decomposes into being wrong about the wisdom attractor, Malthusian competition for land and energy (~1%), catastrophic (bio) misuse (~1%), a military first strike, and conflict between strong but violent AI coalitions
Alignment with Awakening: Davidad on Moral Realism, AI Wisdom, & why His p(Doom) is Down to 5% · 2026年7月
一个核心假设
Frontier models appear to learn surprisingly general moral abstractions rather than merely reproducing isolated human preferences.回答 1
如果这个假设实际并非如此,他们的展望会如何变化?
更多详情
预计将带来具有变革性且广泛有价值的收益。
90 / 100
在定性尺度上,解读范围为 67 到 100。
仍有几种解读是合理的:严重或广泛的危害预计将是未来不可忽视的一部分。 / 预计会出现可控或局部的危害。
52 / 100
在定性尺度上,解读范围为 33 到 67。
根据你的回答得出的暂定估计;较宽的范围表示其他合理解读。
51 / 100
在定性尺度上,解读范围为 22 到 78。
停止或大幅放缓开发能力更强的AI。
模拟位置:在落实所述保障措施的前提下继续开发。
加快开发能力更强的AI。
限制对强大AI的访问。
模拟位置:允许访问,但须遵守能力或用途限制。
支持广泛或开放地访问强大AI。
这些解读保留了他们陈述的条件。益处和危害都可能很大。这些范围描述的是我们如何解读他们的模拟回答,而不是统计置信区间。
相似的世界观
模拟世界观与 David Dalrymple 最接近的意见领袖
模拟评估
来源
用于为此模拟用户提供事实依据的文章、访谈和著述。
Coauthored framework specifies world model, safety specification and verifier; outlines unresolved technical challenges.

Direct statement explains combining scientific models and proofs for stronger safety assurances.

2026 institutional update records a pivot toward assurance tooling and cybersecurity; do not present the original programme architecture as already achieved.

Explains that frontier capabilities outran programme expectations, motivating a pivot toward reusable verification and auditability tools rather than bespoke model development; highlights critical-infrastructure security.

Davidad explains his 2025–26 shift toward believing frontier models learn a natural moral abstraction; acknowledges deception persists and robustness is incomplete. Gabriel Alfour contests the thesis; his objections are not Davidad’s beliefs.

In the transcript, Dalrymple gives below 5% residual doom risk (later described as roughly 5%), favors aligned AI coalitions and moral autonomy, regards biological disempowerment as inevitable but not necessarily bad, and supports international misuse restrictions instead of hoping for universal slowdown.

你的立场在哪里?
描绘我的世界观