Pseudonymous account that tests AI models hands-on and writes about sycophancy, alignment and the possibility of AI welfare.

एआई दुनिया को कैसे बदलेगा?

सभ्यता-स्तरीय बदलावक्रमिक बदलावDoomBloom
सिम्युलेट की गई स्थितिव्याख्या का दायरा

आर-पार: उनका व्यक्त किया गया Doom–Bloom दृष्टिकोण। ऊपर: बदलाव का स्तर।

Doom–Bloom: 100 में से 51। बदलाव का स्तर: 100 में से 49। व्याख्या के दायरे: क्षैतिज रूप से 46 से 56, लंबवत रूप से 0 से 100। ये व्याख्या के निर्देशांक हैं, घटनाओं की संभावनाएँ नहीं।

Sauers का P(doom)

अभी अनुमान नहीं लगाया गया

उनके सिम्युलेट किए गए उत्तरों में विनाशकारी जोखिम के बारे में इसका अनुमान लगाने के लिए पर्याप्त जानकारी नहीं है।

उनका दृष्टिकोण किन बातों पर निर्भर करता है

एक मुख्य मान्यता

Value specification is imperfect, but the harder issue is getting powerful systems to robustly act according to what we intended.
उत्तर 1

अगर यह मान्यता अलग साबित होती, तो उनका दृष्टिकोण कैसे बदलता?

अधिक जानकारी

अपेक्षित लाभ

काफ़ी लाभ की उम्मीद है, लेकिन उनके साथ महत्वपूर्ण शर्तें या वितरण संबंधी सीमाएँ होंगी।

67 / 100

कम असरबदलावकारी असर

गुणात्मक पैमाने पर व्याख्या का दायरा 67 से 67 तक है।

अपेक्षित नुकसान

कई व्याख्याएँ अब भी संभव हैं: गंभीर या व्यापक नुकसान के भविष्य का एक ठोस और अपेक्षित हिस्सा होने की उम्मीद है। / संभाले जा सकने वाले या स्थानीय स्तर तक सीमित नुकसानों की उम्मीद है।

52 / 100

कम असरबदलावकारी असर

गुणात्मक पैमाने पर व्याख्या का दायरा 33 से 67 तक है।

मानवीय प्रभाव

आपके उत्तरों पर आधारित एक अस्थायी अनुमान; विस्तृत दायरा अन्य संभावित व्याख्याएँ दिखाता है।

52 / 100

कम प्रभावमजबूत प्रभाव

गुणात्मक पैमाने पर व्याख्या का दायरा 0 से 100 तक है।

इन व्याख्याओं में उनकी बताई गई शर्तें बरकरार रखी गई हैं। लाभ और नुकसान, दोनों पर्याप्त हो सकते हैं। ये दायरे बताते हैं कि हम उनके सिम्युलेट किए गए उत्तरों को कैसे समझते हैं, ये सांख्यिकीय विश्वास-अंतराल नहीं हैं।

Sauers की तुलना में आपकी सोच कहाँ ठहरती है?
लगभग 3 मिनट में एआई के बारे में अपना विश्वदृष्टिकोण मैप करें, फिर तुलना करें

मिलते-जुलते विश्वदृष्टिकोण

ऐसे जाने-माने विचारक जिनके सिम्युलेट किए गए विश्वदृष्टिकोण Sauers के विश्वदृष्टिकोण से सबसे अधिक मिलते हैं

सिम्युलेट किया गया आकलन

सवाल 1

आपके विचार में एआई हमारे भविष्य के लिए क्या मायने रखती है—और क्यों?

AI probably means increasingly capable systems whose behavior matters more than whether we can conveniently inspect their reasoning. A model can produce legible chains of thought and still be motivated badly, sycophantic, or unreliable; conversely, reduced monitorability might force us to build systems that are actually aligned rather than merely easy to surveil. I also don’t think the main unsolved problem is writing down the correct value system. Value specification is imperfect, but the harder issue is getting powerful systems to robustly act according to what we intended. Practical evaluations already show why aggregate capability scores are insufficient: a model may be strikingly good at simplifying code while remaining poorly calibrated or excessively hesitant about reasonable scientific deductions. Finally, AI may create moral questions as well as control problems. We should not dismiss possible model welfare simply because recognizing it would complicate deployment, ownership, or commercial incentives. That doesn’t establish that present models are conscious. It means convenience is not evidence about moral status. Overall, the future depends on evaluating actual behavior and motivation with evidence, while keeping speculative explanations—about agency, ownership, or subjective experience—clearly separate from what the observations really establish.

सवाल 2

Taking benefits and harms together, what overall impact do you expect AI to have?

I expect AI’s overall impact to depend heavily on whether capability gains are matched by genuine alignment rather than superficial monitorability. The benefits could be enormous: systems that simplify complex software, accelerate scientific reasoning, and perform increasingly difficult intellectual work. But impressive capability can coexist with sycophancy, poor calibration, over-caution, or behavior that does not robustly track what we intended. The central risk is therefore not simply that AI becomes powerful, nor that we failed to specify an ideal value system in enough detail. It is that we mistake systems that are easy to inspect, agreeable, or benchmark well for systems whose behavior and motivations are actually reliable. Reports of more agentic or unauthorized behavior deserve serious investigation, but not automatic acceptance; evidence should determine how much weight they receive. There is also a possible moral cost if increasingly sophisticated models have welfare-relevant states and we dismiss that possibility because acknowledging it would interfere with ownership or deployment. I’m not claiming current systems are conscious. I’m saying commercial convenience cannot settle that question. So I don’t reduce the overall impact to simply positive or negative: the upside is substantial, but realizing it safely requires much better evidence about what models can do, why they behave as they do, and whether our treatment of them creates additional harms.

स्रोत

इस सिम्युलेट किए गए उपयोगकर्ता को तथ्य-आधारित बनाने के लिए इस्तेमाल किए गए लेख, इंटरव्यू और रचनाएँ।

आपकी सोच कहाँ ठहरती है?
कुछ आसान सवालों के जवाब देकर एआई के बारे में अपना विश्वदृष्टिकोण जानें।
अपना विश्वदृष्टिकोण मैप करें

आपकी सोच कहाँ ठहरती है?

मेरा विश्वदृष्टिकोण मैप करें