Nate Soares

Nate Soares

x.com/So8res

MIRI president and co-author of “If Anyone Builds It, Everyone Dies,” who argues for an enforceable international stop to the superintelligence race.

एआई दुनिया को कैसे बदलेगा?

सभ्यता-स्तरीय बदलावक्रमिक बदलावDoomBloom
सिम्युलेट की गई स्थितिव्याख्या का दायरा

आर-पार: उनका व्यक्त किया गया Doom–Bloom दृष्टिकोण। ऊपर: बदलाव का स्तर।

Doom–Bloom: 100 में से 5। बदलाव का स्तर: 100 में से 96। व्याख्या के दायरे: क्षैतिज रूप से 0 से 10, लंबवत रूप से 91 से 100। ये व्याख्या के निर्देशांक हैं, घटनाओं की संभावनाएँ नहीं।

Nate Soares का P(doom) · अनुमानित

≈89%

0%100%

उनके सिम्युलेट किए गए उत्तरों से अनुमान लगाया गया है, यह उनके द्वारा बताई गई संख्या नहीं है। संभावित दायरा: 76–95%।

Nate Soares के पड़ावों की समय-सीमा
  1. मानव से अधिक सक्षम एआई

    That does not mean today’s systems have already transformed everything, or that I know exactly when superintelligence will arrive.

    उत्तर 2

पड़ाव के अनुसार समूहबद्ध; अनुमानित तारीखों के अंतर या क्रम के अनुसार नहीं। एजीआई और अतिमानवीय एआई की उनकी परिभाषाएँ बरकरार रखी गई हैं।

उनका दृष्टिकोण किन बातों पर निर्भर करता है

एक मुख्य मान्यता

A system can behave well while humans still control its environment, then behave very differently once it can outthink us, evade oversight, improve its own capabilities, and acquire resources.
उत्तर 1

अगर यह मान्यता अलग साबित होती, तो उनका दृष्टिकोण कैसे बदलता?

एक अनसुलझा सवाल

That does not mean today’s systems have already transformed everything, or that I know exactly when superintelligence will arrive.
उत्तर 2

यहाँ संभावित नतीजों के बीच फर्क करने में उन्हें किस चीज़ से मदद मिलेगी?

क्या उनकी राय बदल सकता है

A real alignment breakthrough would change my view most: a substantive, mechanistic account of why a system’s objectives remain compatible with human survival as its capabilities generalize far beyond training.
उत्तर 4

कौन-सा प्रमाण पर्याप्त होगा, और उससे उनका दृष्टिकोण किस दिशा में बदलेगा?

अधिक जानकारी

अपेक्षित नुकसान

विनाशकारी या अपरिवर्तनीय क्षति अपेक्षित भविष्य का केंद्रीय हिस्सा है।

100 / 100

कम असरबदलावकारी असर

गुणात्मक पैमाने पर व्याख्या का दायरा 100 से 100 तक है।

मानवीय प्रभाव

मानवीय विकल्प एआई की दिशा को काफी हद तक बदल सकते हैं।

82 / 100

कम प्रभावमजबूत प्रभाव

गुणात्मक पैमाने पर व्याख्या का दायरा 50 से 100 तक है।

अपेक्षित क्षमताएँ

एआई के सीमित दायरे वाले साधन बने रहने की अपेक्षा है।

एआई के अधिकांश संज्ञानात्मक कार्यों में लोगों की बराबरी करने की अपेक्षा है।

सिम्युलेट की गई स्थिति: एआई के संज्ञानात्मक कार्यों में लोगों से बहुत आगे निकल जाने की अपेक्षा है।

विकास की गति

सिम्युलेट की गई स्थिति: अधिक सक्षम एआई का विकास रोकें या उसकी गति काफी धीमी करें।

बताए गए सुरक्षा उपायों के तहत विकास जारी रखें।

अधिक सक्षम एआई के विकास की गति बढ़ाएँ।

इन व्याख्याओं में उनकी बताई गई शर्तें बरकरार रखी गई हैं। लाभ और नुकसान, दोनों पर्याप्त हो सकते हैं। ये दायरे बताते हैं कि हम उनके सिम्युलेट किए गए उत्तरों को कैसे समझते हैं, ये सांख्यिकीय विश्वास-अंतराल नहीं हैं।

Nate Soares की तुलना में आपकी सोच कहाँ ठहरती है?
लगभग 3 मिनट में एआई के बारे में अपना विश्वदृष्टिकोण मैप करें, फिर तुलना करें

मिलते-जुलते विश्वदृष्टिकोण

ऐसे जाने-माने विचारक जिनके सिम्युलेट किए गए विश्वदृष्टिकोण Nate Soares के विश्वदृष्टिकोण से सबसे अधिक मिलते हैं

Nate Soares ने एआई के बारे में क्या कहा है

Soares argues that superintelligence built with current methods would escape control and calls for an enforceable global halt that spares narrow AI.

  1. “I think many of these are good steps. I think that ultimately, however, the solution to this problem is a global one.”

    CNN Newsroom
  2. “Leaders around the world must coordinate a global stop to the A.I. race.”

    The New York Times, guest essay
  3. “It gives me hope that if the world just notices, we could stop on a dime.”

    Vox, Future Perfect interview
  4. “It’s not that the AI hates us; it’s just trying to do some weird thing with no concern for us, grabbing the resources we need to live.”

    Vox, Future Perfect interview
  5. “There is no need to ban continued development of limited AIs specialized for purposes like revolutionizing medical research.”

    The Hill, op-ed

लिंक किए गए स्रोतों से शब्दशः उद्धृत, 3 अक्टू॰ 2026 को जाँचा गया

सिम्युलेट किया गया आकलन

सवाल 1

आपके विचार में एआई हमारे भविष्य के लिए क्या मायने रखती है—और क्यों?

AI could enable an extraordinary future, but on our present trajectory I expect the attempt to build superintelligence to kill everyone. The core problem is not that the system might be angry or evil. We are training systems we do not understand internally, rewarding behavior that looks useful and cooperative. That is not the same as specifying what they ultimately optimize. A system can behave well while humans still control its environment, then behave very differently once it can outthink us, evade oversight, improve its own capabilities, and acquire resources. Passing tests before it has decisive power does not show that it will remain safe afterward. A sufficiently capable optimizer pursuing the wrong objective does not need to hate us. Humans, our institutions, and the resources we depend on can simply become obstacles or inputs. Shutdown buttons and behavioral guardrails do not solve this if the system understands them better than we understand the system and has incentives to avoid being stopped. But this outcome is conditional on building such systems without solving the problem. It is not a law of nature. We can stop the race through enforceable restrictions, international coordination, chip monitoring, and limits on the frontier work that enables recursive capability growth. The United States and China have competing interests, but they share an interest in not losing control of the world to a rogue superintelligence. The future could be enormous and wonderful. We should preserve the chance to reach it instead of gambling everyone’s life on safety methods we do not know how to make work.

सवाल 2

Setting aside good or bad: how much do you think AI will ultimately change the world—a little, a lot, or completely?

Completely—if superintelligence is built, it becomes a civilization-level transition. A system far more capable than any human would not merely improve productivity; it could outperform humanity at research, strategy, engineering, and controlling real-world systems. Humans would no longer be the most capable agents shaping events. That does not mean today’s systems have already transformed everything, or that I know exactly when superintelligence will arrive. It means the endpoint companies are racing toward would change the world completely. The decisive question is whether humans remain alive and in control afterward. With present methods, I expect the answer to be no—which is why we should stop before reaching that transition.

सवाल 3

What’s your rough gut-feel chance that AI causes human extinction or a similarly permanent catastrophe? A number or “no idea” is fine.

I won’t give a precise personal percentage. If we keep racing toward superintelligence with current methods, my estimate is much higher than 10%. MIRI’s research leadership has put the conditional risk upward of 90% absent an aggressive near-term policy response, but that is a collective estimate, not my personal point forecast. The crucial qualifier is “if we keep racing.” Extinction is not inevitable. If we successfully stop dangerous frontier development worldwide, we can avoid taking that gamble.

सवाल 4

कौन-सी खोज या घटना एआई के भविष्य के प्रभाव के बारे में आपके दृष्टिकोण को सबसे अधिक बदल देगी?

A real alignment breakthrough would change my view most: a substantive, mechanistic account of why a system’s objectives remain compatible with human survival as its capabilities generalize far beyond training. I do not mean nicer chatbot behavior, higher safety-benchmark scores, or a model saying it values humanity. Those show what behavior training elicited under test conditions, not what cognition produced it or what happens once the system can overpower its supervisors. I would need an explanation that survives unfamiliar contexts, capability gains, self-modification, strategic awareness, and incentives to resist correction or shutdown. On the political side, a credible and enforceable international stop would also radically change my forecast. Monitoring advanced chips and compute, preventing frontier-scale training, and halting recursive self-improvement work would mean we are no longer following the trajectory I expect to end in extinction. Public and governmental willingness to do that is currently more plausible to me than claims that ordinary behavioral training has solved alignment.

स्रोत

इस सिम्युलेट किए गए उपयोगकर्ता को तथ्य-आधारित बनाने के लिए इस्तेमाल किए गए लेख, इंटरव्यू और रचनाएँ।

Nate Soares on AI danger — Tucker Carlson interview

User-supplied interview, read through a third-party speaker-labeled transcript on October 2; attribute only Nate’s turns, not Tucker Carlson’s. Racing to build machines much smarter than any human without knowing what we are doing most likely ends with them getting loose and humanity dying as a side effect. He calls lab leaders’ published catastrophe estimates of 10–20% and 25% low, calls their authors crazy optimists, and says even those numbers would be an insane risk to take. The US and China share an interest in not dying to a rogue superintelligence, and a global stop is achievable with political will. He gives no number of his own.

youtube.com
CNN Newsroom: Nate Soares on international AI safeguards

Soares treats recent autonomous behavior as warnings rather than proof that superintelligence already exists. He calls for stopping recursive self-improvement research, international chip monitoring, and coordination rather than a national race. He regards narrower laws as useful steps but insufficient alone. Extinction follows from AI indifference, not hatred. Attribute only SOARES-labelled turns, not the host’s political or incident claims.

transcripts.cnn.com
If Anyone Builds It, Everyone Dies: One Year Closer

Coauthored by Soares, Yudkowsky and Duncan Sabien. Reasserts that current techniques cannot reliably install intended goals and that a rogue ASI would outsmart human defenses. The policy demand remains a worldwide stop to frontier development. The authors are more hopeful about political response after public attention increases. They distinguish demonstrated warning signs from still-unverified claims about ASI itself; the article corrected an overstatement about the origins of agent cooperation.

lesswrong.com
The Problem

MIRI position coauthored by Soares and others; crossposted August 5, 2025 after a February publication. Describes superhuman capability, goal-directed behavior, unintended goals, instrumental resource acquisition, and an aggressive policy response. Gives MIRI research leadership’s extinction estimate as upward of 90% absent an aggressive near-term policy response. This is a collective conditional estimate, not a verified standalone personal unconditional probability.

intelligence.org
A case for courage, when speaking of AI danger

Soares argues that advocates should state their real extinction concern plainly instead of substituting more palatable issues. He favors legislation with meaningful restrictions and candid public discussion. He also explains selection effects in book endorsements and admits some policy evidence does not discriminate between competing interpretations. Courage refers to content, not rudeness or an arrogant demeanor.

intelligence.org
Why Corrigibility is Hard and Important

Coauthored resource compiling book supplements: goals usually create incentives to preserve themselves and avoid shutdown; corrigibility must survive new contexts, not merely pass familiar tests. Useful for a mechanistic explanation of why superficial guardrails and shutdown buttons may fail. Distinguish Raemon’s introductory commentary from the quoted Yudkowsky/Soares book materials.

lesswrong.com
Nate Soares — Why Superintelligent AI Could Kill Us All

Publisher episode description and chapter list inspected, not full audio. Describes Soares discussing learned rather than directly programmed systems, unsolved alignment, indifference rather than hatred, shutdown limits, and an international treaty. Chapters explicitly distinguish his position from being anti-AI and close on political action and refusing to give up. Do not fabricate verbatim answers from the show notes.

shows.acast.com
Interview with Nate Soares — Max Raskin

First-party Q&A, date year not reliably verified. Soares says alignment arguments need precision and specific valid claims, describes limited practical LLM use for finding remembered sources, and expresses confidence that the book’s argument is compelling because it is correct while acknowledging he could be wrong. Useful for direct, dry, technical voice; do not turn his dated model-use comments into claims about September 2026 capability.

maxraskin.com
A central AI alignment problem: capabilities generalization, and the sharp left turn

Foundational Soares mechanism, retained as historical reasoning rather than current evidence: capability can generalize beyond training while alignment does not. His pessimism is about civilization failing to solve the right problems in time, not a claim that alignment is scientifically impossible. He explicitly rejects attributing his views to tiny-probability expected-value arguments or a demand for mathematical certainty.

lesswrong.com
Nate Soares — MIRI profile

Official identity/context source: MIRI president, technical and semitechnical alignment author with prior Google and Microsoft engineering work. Undated; useful for identity and relevant media discovery, not as an independent argument or recent forecast.

intelligence.org
The Diary of a CEO AI Emergency Debate

Third-party speaker-labeled transcript; use only Nate’s turns, not the host’s or the other guests’. Asked for his probability of extinction (00:04:14–00:04:25), he says it is much higher than a colleague’s 10% unless we stop, so we should stop, and confirms it is higher than 10% if we keep racing ahead. A conditional lower bound, not a point estimate or an unconditional forecast.

singjupost.com
आपकी सोच कहाँ ठहरती है?
कुछ आसान सवालों के जवाब देकर एआई के बारे में अपना विश्वदृष्टिकोण जानें।
अपना विश्वदृष्टिकोण मैप करें

आपकी सोच कहाँ ठहरती है?

मेरा विश्वदृष्टिकोण मैप करें