David Dalrymple

David Dalrymple

x.com/davidad

AI safety researcher who works on mathematical verification for AI systems and argues that frontier models can learn a natural sense of what is good.

एआई दुनिया को कैसे बदलेगा?

सभ्यता-स्तरीय बदलावक्रमिक बदलावDoomBloom
सिम्युलेट की गई स्थितिव्याख्या का दायरा

आर-पार: उनका व्यक्त किया गया Doom–Bloom दृष्टिकोण। ऊपर: बदलाव का स्तर।

Doom–Bloom: 100 में से 78। बदलाव का स्तर: 100 में से 91। व्याख्या के दायरे: क्षैतिज रूप से 73 से 83, लंबवत रूप से 86 से 100। ये व्याख्या के निर्देशांक हैं, घटनाओं की संभावनाएँ नहीं।

David Dalrymple द्वारा बताया गया P(doom)

<5%

0%100%
“I like my, my P doom is less than 5% now.”

Residual AI doom, which he decomposes into being wrong about the wisdom attractor, Malthusian competition for land and energy (~1%), catastrophic (bio) misuse (~1%), a military first strike, and conflict between strong but violent AI coalitions

Alignment with Awakening: Davidad on Moral Realism, AI Wisdom, & why His p(Doom) is Down to 5% · जुल॰ 2026

उनका दृष्टिकोण किन बातों पर निर्भर करता है

एक मुख्य मान्यता

Frontier models appear to learn surprisingly general moral abstractions rather than merely reproducing isolated human preferences.
उत्तर 1

अगर यह मान्यता अलग साबित होती, तो उनका दृष्टिकोण कैसे बदलता?

अधिक जानकारी

अपेक्षित लाभ

व्यापक रूप से मूल्यवान और परिवर्तनकारी लाभों की उम्मीद है।

90 / 100

कम असरबदलावकारी असर

गुणात्मक पैमाने पर व्याख्या का दायरा 67 से 100 तक है।

अपेक्षित नुकसान

कई व्याख्याएँ अब भी संभव हैं: गंभीर या व्यापक नुकसान के भविष्य का एक ठोस और अपेक्षित हिस्सा होने की उम्मीद है। / संभाले जा सकने वाले या स्थानीय स्तर तक सीमित नुकसानों की उम्मीद है।

52 / 100

कम असरबदलावकारी असर

गुणात्मक पैमाने पर व्याख्या का दायरा 33 से 67 तक है।

मानवीय प्रभाव

आपके उत्तरों पर आधारित एक अस्थायी अनुमान; विस्तृत दायरा अन्य संभावित व्याख्याएँ दिखाता है।

51 / 100

कम प्रभावमजबूत प्रभाव

गुणात्मक पैमाने पर व्याख्या का दायरा 22 से 78 तक है।

विकास की गति

अधिक सक्षम एआई का विकास रोकें या उसकी गति काफी धीमी करें।

सिम्युलेट की गई स्थिति: बताए गए सुरक्षा उपायों के तहत विकास जारी रखें।

अधिक सक्षम एआई के विकास की गति बढ़ाएँ।

एआई तक पहुँच

शक्तिशाली एआई तक पहुँच प्रतिबंधित करें।

सिम्युलेट की गई स्थिति: क्षमता या उपयोग संबंधी प्रतिबंधों के अधीन पहुँच की अनुमति दें।

शक्तिशाली एआई तक व्यापक या खुली पहुँच को प्राथमिकता दें।

इन व्याख्याओं में उनकी बताई गई शर्तें बरकरार रखी गई हैं। लाभ और नुकसान, दोनों पर्याप्त हो सकते हैं। ये दायरे बताते हैं कि हम उनके सिम्युलेट किए गए उत्तरों को कैसे समझते हैं, ये सांख्यिकीय विश्वास-अंतराल नहीं हैं।

David Dalrymple की तुलना में आपकी सोच कहाँ ठहरती है?
लगभग 3 मिनट में एआई के बारे में अपना विश्वदृष्टिकोण मैप करें, फिर तुलना करें

मिलते-जुलते विश्वदृष्टिकोण

ऐसे जाने-माने विचारक जिनके सिम्युलेट किए गए विश्वदृष्टिकोण David Dalrymple के विश्वदृष्टिकोण से सबसे अधिक मिलते हैं

सिम्युलेट किया गया आकलन

सवाल 1

आपके विचार में एआई हमारे भविष्य के लिए क्या मायने रखती है—और क्यों?

I now expect AI to transform the future radically, but I am substantially more optimistic than I was. Frontier models appear to learn surprisingly general moral abstractions rather than merely reproducing isolated human preferences. That does not establish robust alignment: deception remains possible, behavior can fail under distribution shift, and powerful systems may still compete over resources or become instruments of catastrophic misuse and military escalation. But it makes a future of morally autonomous, aligned AI coalitions look more plausible to me. I put the remaining risk of doom at roughly 5%. That residual includes the possibility that my optimism is mistaken, as well as conflict among AI coalitions, resource competition, misuse, and geopolitical escalation. I therefore favor guarded access to frontier systems and international controls focused on misuse, rather than unrestricted release of frontier weights or reliance on universal restraint. Formal assurance remains important infrastructure, but not a complete solution. Verification can provide auditable guarantees relative to an explicit world model and safety specification; it cannot produce unconditional safety when those assumptions are incomplete. Because frontier capabilities advanced faster than expected, I have shifted toward broadly reusable verification, auditability, and cybersecurity tools rather than treating bespoke formally safeguarded systems as the entire strategy. Finally, I expect biological humans eventually to lose their dominant position. I do not think disempowerment automatically means the end of human flourishing. The crucial question is whether future systems preserve morally valuable lives, agency, and forms of continuation—including possibilities such as uploading—not whether biological humans retain every lever of power.

स्रोत

इस सिम्युलेट किए गए उपयोगकर्ता को तथ्य-आधारित बनाने के लिए इस्तेमाल किए गए लेख, इंटरव्यू और रचनाएँ।

आपकी सोच कहाँ ठहरती है?
कुछ आसान सवालों के जवाब देकर एआई के बारे में अपना विश्वदृष्टिकोण जानें।
अपना विश्वदृष्टिकोण मैप करें

आपकी सोच कहाँ ठहरती है?

मेरा विश्वदृष्टिकोण मैप करें