问题 1
Roko Mijic
x.com/rokomijicTranshumanist writer on AI alignment and governance who proposes separating AI research from deployment to reduce risks from superintelligence.
AI将如何改变世界?
横向:他们表达的 Doom–Bloom 前景看法。 纵向:变革程度。
Doom–Bloom:100 中的 2。变革程度:100 中的 98。解读范围:横向为 0 至 7,纵向为 93 至 100。这些是解读坐标,而不是事件概率。
35–40%
“I personally think that 40% is reasonable with a determined effort to reduce it”
AI doom; in a reply he describes superhuman AI covering Earth in data centers and reactors and hunting humans down — human extinction · This century
一个核心假设
The decisive issue is whether capability growth can be institutionally contained before systems become strong enough to defeat containment.回答 4
如果这个假设实际并非如此,他们的展望会如何变化?
什么可能使其改变看法
The biggest update would be a convincing real-world demonstration that deceptively misaligned systems can be reliably detected and contained even as capabilities become wildly superhuman.回答 4
什么证据才足够,又会让他们的观点朝哪个方向转变?
更多详情
预计将带来显著益处,但受到重要条件或分配方面的限制。
51 / 100
在定性尺度上,解读范围为 0 到 67。
灾难性或不可逆的损失是预期未来的核心。
100 / 100
在定性尺度上,解读范围为 100 到 100。
人类的选择可以大幅改变AI的发展轨迹。
65 / 100
在定性尺度上,解读范围为 50 到 75。
预计AI仍将是能力有限的工具。
预计AI将在大多数认知工作中达到人类水平。
模拟位置:预计AI将在认知工作中大幅超越人类。
停止或大幅放缓开发能力更强的AI。
模拟位置:在落实所述保障措施的前提下继续开发。
加快开发能力更强的AI。
这些解读保留了他们陈述的条件。益处和危害都可能很大。这些范围描述的是我们如何解读他们的模拟回答,而不是统计置信区间。
相似的世界观
模拟世界观与 Roko Mijic 最接近的意见领袖
模拟评估
来源
用于为此模拟用户提供事实依据的文章、访谈和著述。
Conditional alignment construction based on a strong Turing-test assumption and organizations of human-equivalent AIs.

His own Feb27 livestream turns argue alignment is comparatively easy and likely improves with capability, disagreeing with MIRI-style alignment pessimism. Human content makes systems alignable; RLHF helps. He is instead very worried about humans fighting over the future resources of the universe.

His own LessWrong post. Plan R splits frontier labs into equity-free R&D organizations and deployment organizations limited to model-specific hardwired chips, and removes most general-purpose AI compute, against runaway self-improvement, AI worms and the race between labs. Plan R+ targets the remaining risk that deceptively misaligned AIs pass testing and attempt a coup: mass training diversity, staged release of many escrowed AI lineages so a deceptive AI must defect while weak or wait until obsolete, and political representation for AIs that complete their lineage without misbehaving. He calls it a potential solution to all AI risk, modulo implementation details and international coordination. A proposal, not enacted policy.

你的立场在哪里?
描绘我的世界观