質問1
Rob Bensinger
x.com/robbensingerMIRI writer who argues superhuman AI built with current methods would be too dangerous and calls for an international halt to the race to build it.
AIは世界をどのように変えるでしょうか?
横軸:彼が表明したDoom–Bloomの見通し。 縦軸:変革の規模。
Doom–Bloom:100点中4。変革の規模:100点中97。解釈範囲:横方向は0から9、縦方向は92から100。これらは解釈上の座標であり、事象の確率ではありません。
≈72%
本人が示した数値ではなく、シミュレーションされた本人の回答から推定したものです。 妥当と考えられる範囲:57–84%。
中心的な前提
As systems become more capable and agentic—better at planning, persisting, and routing around obstacles—the cost of getting those goals slightly wrong becomes catastrophic.回答1
この前提が実際には異なると判明した場合、彼の見通しはどう変わりますか?
考えを変え得るもの
The biggest update would be a real, legible theory of alignment: one that lets us understand and reliably control the internal goals of systems smarter than us, rather than merely patching their visible behavior.回答4
どのような証拠なら十分で、それによって彼の見解はどちらの方向に変わりますか?
詳細
複数の解釈が依然として妥当です:高度なAIが登場しても、良い影響はほとんどないと予想されています。 / 大きな恩恵が予想されていますが、重要な条件や分配上の制約があります。 / 恩恵は限定的、または狭い範囲にしか行き渡らないと予想されています。
34 / 100
質的尺度での解釈範囲は0から67です。
破局的または不可逆的な喪失が、予想される将来の中心となっています。
100 / 100
質的尺度での解釈範囲は100から100です。
人間の選択によって、AIの軌道を大幅に変えることができます。
69 / 100
質的尺度での解釈範囲は50から100です。
AIは、限定的なツールにとどまると予想されています。
AIは、ほとんどの認知作業において人間と同等になると予想されています。
シミュレーション上の位置:AIは、認知作業全般において人間を大幅に上回ると予想されています。
シミュレーション上の位置:より高性能なAIの開発を停止するか、大幅に減速させます。
明示された安全対策の下で開発を継続します。
より高性能なAIの開発を加速させます。
これらの解釈では、彼が示した条件が維持されています。恩恵と害は、どちらも大きくなり得ます。この範囲は、統計的な信頼区間ではなく、彼のシミュレーションされた回答をどのように読み取ったかを示すものです。
似ている世界観
シミュレーションされた世界観がRob Bensingerの世界観に最も近いオピニオンリーダー
シミュレーション評価
出典
このシミュレーション対象者の根拠として使用された記事、インタビュー、著作です。
Full post inspected; its subtitle says it was written August 7 and published later. Explains the world’s slow response through machine learning’s trial-and-error culture, difficulty reckoning emotionally with a new kind of entity, social risk, an online “irony mandate,” and too few senior people taking engineering ownership of the danger. Says the window for international response is plausibly closing soon, if not already closed. Frames these failures as a choice that can be reversed, not destiny. Quoted remarks by Soares, Sam Harris and Joshua Achiam are theirs.

Full comment inspected via the LessWrong API. Argues that Anthropic’s and Dario Amodei’s visible messaging leaves a large candor gap relative to what many of their own researchers believe, and criticizes Anthropic for opposing US–China coordination and pursuing recursive self-improvement. He calls OpenPhil’s bet on OpenAI a disaster, while noting he had said EA’s net effect on x-risk was probably positive but highly uncertain. He says Anthropic may or may not be slightly better than OpenAI. Quoted statements by Greenblatt, Buck and others are theirs.

Full post inspected. Proposes a simultaneous, US-brokered international halt on the race to superintelligence, enforced through the concentrated chip supply chain with monitoring and possibly kill switches. The ban would last until it is clear we can build superintelligence safely, which could mean decades, and would leave existing AI and inference largely untouched. Rebuts concerns about cost, totalitarianism, defectors and China, and argues a unilateral US halt would be counterproductive. Cites Jan Leike’s 10–90% and Dario Amodei’s 25% as others’ estimates, not his own.

Older context, with the opening sections and takeoff discussion inspected. He argues that Will MacAskill’s optimism rests on a fragile conjunction of premises, so a double-digit chance of ruin remains even if each premise looks plausible. He also argues that soft, continuous takeoff would not meaningfully improve survival odds, and that good behavior from weak AIs does not show a superintelligence would be aligned. He writes partly as a MIRI insider defending the book and quotes Yudkowsky. Newer 2026 sources take precedence for current policy specifics.

Older institutional context; the byline and opening section were inspected. States MIRI’s view that building superintelligent AI with anything like current understanding or methods has human extinction as its expected outcome, and calls for governments to halt development. Use it as the shared MIRI frame Rob helped write, not as his individual phrasing. Its numerical extinction estimate is attributed to MIRI research leadership and is not his personal P(doom).

あなたはどの位置でしょうか?
自分の世界観をマッピングする