質問1
Roko Mijic
x.com/rokomijicTranshumanist writer on AI alignment and governance who proposes separating AI research from deployment to reduce risks from superintelligence.
AIは世界をどのように変えるでしょうか?
横軸:その人が表明したDoom–Bloomの見通し。 縦軸:変革の規模。
Doom–Bloom:100点中2。変革の規模:100点中98。解釈範囲:横方向は0から7、縦方向は93から100。これらは解釈上の座標であり、事象の確率ではありません。
35–40%
“I personally think that 40% is reasonable with a determined effort to reduce it”
AI doom; in a reply he describes superhuman AI covering Earth in data centers and reactors and hunting humans down — human extinction · This century
中心的な前提
The decisive issue is whether capability growth can be institutionally contained before systems become strong enough to defeat containment.回答4
この前提が実際には異なると判明した場合、その人の見通しはどう変わりますか?
考えを変え得るもの
The biggest update would be a convincing real-world demonstration that deceptively misaligned systems can be reliably detected and contained even as capabilities become wildly superhuman.回答4
どのような証拠なら十分で、それによってその人の見解はどちらの方向に変わりますか?
詳細
大きな恩恵が予想されていますが、重要な条件や分配上の制約があります。
51 / 100
質的尺度での解釈範囲は0から67です。
破局的または不可逆的な喪失が、予想される将来の中心となっています。
100 / 100
質的尺度での解釈範囲は100から100です。
人間の選択によって、AIの軌道を大幅に変えることができます。
65 / 100
質的尺度での解釈範囲は50から75です。
AIは、限定的なツールにとどまると予想されています。
AIは、ほとんどの認知作業において人間と同等になると予想されています。
シミュレーション上の位置:AIは、認知作業全般において人間を大幅に上回ると予想されています。
より高性能なAIの開発を停止するか、大幅に減速させます。
シミュレーション上の位置:明示された安全対策の下で開発を継続します。
より高性能なAIの開発を加速させます。
これらの解釈では、その人が示した条件が維持されています。恩恵と害は、どちらも大きくなり得ます。この範囲は、統計的な信頼区間ではなく、その人のシミュレーションされた回答をどのように読み取ったかを示すものです。
似ている世界観
シミュレーションされた世界観がRoko Mijicの世界観に最も近いオピニオンリーダー
シミュレーション評価
出典
このシミュレーション対象者の根拠として使用された記事、インタビュー、著作です。
Conditional alignment construction based on a strong Turing-test assumption and organizations of human-equivalent AIs.

His own Feb27 livestream turns argue alignment is comparatively easy and likely improves with capability, disagreeing with MIRI-style alignment pessimism. Human content makes systems alignable; RLHF helps. He is instead very worried about humans fighting over the future resources of the universe.

His own LessWrong post. Plan R splits frontier labs into equity-free R&D organizations and deployment organizations limited to model-specific hardwired chips, and removes most general-purpose AI compute, against runaway self-improvement, AI worms and the race between labs. Plan R+ targets the remaining risk that deceptively misaligned AIs pass testing and attempt a coup: mass training diversity, staged release of many escrowed AI lineages so a deceptive AI must defect while weak or wait until obsolete, and political representation for AIs that complete their lineage without misbehaving. He calls it a potential solution to all AI risk, modulo implementation details and international coordination. A proposal, not enacted policy.

あなたはどの位置でしょうか?
自分の世界観をマッピングする