सवाल 1
David Dalrymple
x.com/davidadAI safety researcher who works on mathematical verification for AI systems and argues that frontier models can learn a natural sense of what is good.
एआई दुनिया को कैसे बदलेगा?
आर-पार: उनका व्यक्त किया गया Doom–Bloom दृष्टिकोण। ऊपर: बदलाव का स्तर।
Doom–Bloom: 100 में से 78। बदलाव का स्तर: 100 में से 91। व्याख्या के दायरे: क्षैतिज रूप से 73 से 83, लंबवत रूप से 86 से 100। ये व्याख्या के निर्देशांक हैं, घटनाओं की संभावनाएँ नहीं।
<5%
“I like my, my P doom is less than 5% now.”
Residual AI doom, which he decomposes into being wrong about the wisdom attractor, Malthusian competition for land and energy (~1%), catastrophic (bio) misuse (~1%), a military first strike, and conflict between strong but violent AI coalitions
Alignment with Awakening: Davidad on Moral Realism, AI Wisdom, & why His p(Doom) is Down to 5% · जुल॰ 2026
एक मुख्य मान्यता
Frontier models appear to learn surprisingly general moral abstractions rather than merely reproducing isolated human preferences.उत्तर 1
अगर यह मान्यता अलग साबित होती, तो उनका दृष्टिकोण कैसे बदलता?
अधिक जानकारी
व्यापक रूप से मूल्यवान और परिवर्तनकारी लाभों की उम्मीद है।
90 / 100
गुणात्मक पैमाने पर व्याख्या का दायरा 67 से 100 तक है।
कई व्याख्याएँ अब भी संभव हैं: गंभीर या व्यापक नुकसान के भविष्य का एक ठोस और अपेक्षित हिस्सा होने की उम्मीद है। / संभाले जा सकने वाले या स्थानीय स्तर तक सीमित नुकसानों की उम्मीद है।
52 / 100
गुणात्मक पैमाने पर व्याख्या का दायरा 33 से 67 तक है।
आपके उत्तरों पर आधारित एक अस्थायी अनुमान; विस्तृत दायरा अन्य संभावित व्याख्याएँ दिखाता है।
51 / 100
गुणात्मक पैमाने पर व्याख्या का दायरा 22 से 78 तक है।
अधिक सक्षम एआई का विकास रोकें या उसकी गति काफी धीमी करें।
सिम्युलेट की गई स्थिति: बताए गए सुरक्षा उपायों के तहत विकास जारी रखें।
अधिक सक्षम एआई के विकास की गति बढ़ाएँ।
शक्तिशाली एआई तक पहुँच प्रतिबंधित करें।
सिम्युलेट की गई स्थिति: क्षमता या उपयोग संबंधी प्रतिबंधों के अधीन पहुँच की अनुमति दें।
शक्तिशाली एआई तक व्यापक या खुली पहुँच को प्राथमिकता दें।
इन व्याख्याओं में उनकी बताई गई शर्तें बरकरार रखी गई हैं। लाभ और नुकसान, दोनों पर्याप्त हो सकते हैं। ये दायरे बताते हैं कि हम उनके सिम्युलेट किए गए उत्तरों को कैसे समझते हैं, ये सांख्यिकीय विश्वास-अंतराल नहीं हैं।
मिलते-जुलते विश्वदृष्टिकोण
ऐसे जाने-माने विचारक जिनके सिम्युलेट किए गए विश्वदृष्टिकोण David Dalrymple के विश्वदृष्टिकोण से सबसे अधिक मिलते हैं
सिम्युलेट किया गया आकलन
स्रोत
इस सिम्युलेट किए गए उपयोगकर्ता को तथ्य-आधारित बनाने के लिए इस्तेमाल किए गए लेख, इंटरव्यू और रचनाएँ।
Coauthored framework specifies world model, safety specification and verifier; outlines unresolved technical challenges.

Direct statement explains combining scientific models and proofs for stronger safety assurances.

2026 institutional update records a pivot toward assurance tooling and cybersecurity; do not present the original programme architecture as already achieved.

Explains that frontier capabilities outran programme expectations, motivating a pivot toward reusable verification and auditability tools rather than bespoke model development; highlights critical-infrastructure security.

Davidad explains his 2025–26 shift toward believing frontier models learn a natural moral abstraction; acknowledges deception persists and robustness is incomplete. Gabriel Alfour contests the thesis; his objections are not Davidad’s beliefs.

In the transcript, Dalrymple gives below 5% residual doom risk (later described as roughly 5%), favors aligned AI coalitions and moral autonomy, regards biological disempowerment as inevitable but not necessarily bad, and supports international misuse restrictions instead of hoping for universal slowdown.

आपकी सोच कहाँ ठहरती है?
मेरा विश्वदृष्टिकोण मैप करें