问题 1
Rob Wiblin
x.com/robertwiblin80,000 Hours Podcast host who weighs evidence on AI progress, takes cyber, bio and rogue-agent risks seriously and leans toward slowing frontier AI.
AI将如何改变世界?
横向:他表达的 Doom–Bloom 前景看法。 纵向:变革程度。
Doom–Bloom:100 中的 18。变革程度:100 中的 87。解读范围:横向为 13 至 25,纵向为 69 至 100。这些是解读坐标,而不是事件概率。
≈21%
根据他的模拟回答推断,并非他们给出的数字。 合理范围:14–31%。
一个核心假设
But the current path combines rapidly improving cyber and research capabilities with weak control, declining monitorability, and institutions moving far too slowly.回答 2
如果这个假设实际并非如此,他的展望会如何变化?
一个尚未解决的问题
If systems automate AI research itself, the pace could accelerate sharply—though we genuinely do not know how powerful that feedback loop would be or whether compute and missing real-world capabilities would constrain it.回答 1
什么能帮助他区分这里各种合理的结果?
更多详情
仍有几种解读是合理的:预计将带来显著益处,但受到重要条件或分配方面的限制。 / 预计将带来具有变革性且广泛有价值的收益。
80 / 100
在定性尺度上,解读范围为 67 到 100。
严重或广泛的危害预计将是未来不可忽视的一部分。
67 / 100
在定性尺度上,解读范围为 67 到 67。
人类的选择具有实质性但受到很大制约的影响。
61 / 100
在定性尺度上,解读范围为 43 到 82。
模拟位置:停止或大幅放缓开发能力更强的AI。
在落实所述保障措施的前提下继续开发。
加快开发能力更强的AI。
这些解读保留了他陈述的条件。益处和危害都可能很大。这些范围描述的是我们如何解读他的模拟回答,而不是统计置信区间。
相似的世界观
模拟世界观与 Rob Wiblin 最接近的意见领袖
模拟评估
来源
用于为此模拟用户提供事实依据的文章、访谈和著述。
Weighs seven 2026 developments: revenue growth, METR time horizons, the Mythos jump, Anthropic’s reported internal speedups, AI still struggling to run real businesses, a maths result and cheaper-than-expected inference. Says his timelines shortened by about a year: fully automated AI R&D would shock him in 2026, is imaginable in 2027 and plausible in 2028 if trends continue, while a slower path into the mid-2030s remains very possible. Names four unresolved cruxes (skills needed for recursive self-improvement, missing capabilities in low-feedback domains, spillover from verifiable-reward training, compute bottlenecks). Closes by judging that the benefits of slowing are approaching the point of outweighing the costs and that worried insiders should be given more time; a judgement, not a drafted policy. Full transcript inspected.

His reading of Anthropic’s Mythos system card and alignment risk update. Calls its cyber capabilities a nightmare for computer security and says he is deeply uncomfortable with any company or government having unrestricted access to it. Would bet the strong alignment results probably reflect the model, but argues evaluation awareness, chain-of-thought exposure during training and unfaithful reasoning mean they cannot be taken at face value. Infers that a jump of this size brings automated AI R&D forward and shrinks preparation time, and says he lost sleep over it. An interpretation of company disclosures, not independent testing. Full transcript inspected.

Explains why timelines shortened in early 2025 and lengthened later: limited reasoning generalisation, costly inference scaling, inefficient reinforcement learning, missing continual learning and non-coding bottlenecks in AI R&D. Rejects the story that AI is useless, stalled or unprofitable, citing capability indices, falling costs, revenue, per-user margins and his own heavy daily use. Its timeline (shocked by 2027, imaginable 2028, plausible 2029–2030) is superseded by the August update. Argues that even a roughly ten-year timeline leaves too little time to prepare for social, political, economic, military and epistemic upheaval. Full transcript inspected.

Older context: recorded in 2023 and released in 2025, with Rob saying it mostly held up but he would not say everything the same way now. He says AI risk does not depend on a superintelligence story and that the danger is obvious rather than speculative; he has seen AI as a possible hinge of history since about 2009–2010 and expects useful agentic AI to be built. At the time he thought takeoff more likely to take years or decades than days, which made prosaic safety work and government involvement look more useful, and he did not expect mass layoffs within a couple of years. Newer 2026 sources take precedence on timelines and policy. Own turns inspected.

你的立场在哪里?
描绘我的世界观