Dwarkesh Patel

Dwarkesh Patel

x.com/dwarkesh_sp

Podcast host and essayist who examines how AI systems learn, whether AI research can be automated and the economic and control questions that follow.

AI将如何改变世界?

文明层面的变革渐进式变化DoomBloom
模拟位置解读范围

横向:他表达的 Doom–Bloom 前景看法。 纵向:变革程度。

Doom–Bloom:100 中的 43。变革程度:100 中的 82。解读范围:横向为 38 至 75,纵向为 73 至 100。这些是解读坐标,而不是事件概率。

Dwarkesh Patel的 P(doom) · 推断

≈20%

0%100%

根据他的模拟回答推断,并非他们给出的数字。 合理范围:13–30%。

他的展望取决于什么

一个核心假设

The hinge is whether systems can learn from messy work experience and automate AI research.
回答 2

如果这个假设实际并非如此,他的展望会如何变化?

一个尚未解决的问题

I used to be more skeptical of rapid self-improvement; I now think a large speedup is plausible enough that we have to take it seriously, without pretending we know its timing.
回答 1

什么能帮助他区分这里各种合理的结果?

更多详情

预期益处

预计将带来显著益处,但受到重要条件或分配方面的限制。

74 / 100

影响小变革性影响

在定性尺度上,解读范围为 67 到 100。

预期危害

严重或广泛的危害预计将是未来不可忽视的一部分。

68 / 100

影响小变革性影响

在定性尺度上,解读范围为 67 到 67。

人类影响力

人类的选择具有实质性但受到很大制约的影响。

53 / 100

影响力小影响力强

在定性尺度上,解读范围为 44 到 81。

预期能力

预计AI仍将是能力有限的工具。

预计AI将在大多数认知工作中达到人类水平。

模拟位置:预计AI将在认知工作中大幅超越人类。

这些解读保留了他陈述的条件。益处和危害都可能很大。这些范围描述的是我们如何解读他的模拟回答,而不是统计置信区间。

与Dwarkesh Patel相比,你的立场在哪里?
用大约3分钟描绘你自己的AI世界观,然后进行比较

相似的世界观

模拟世界观与 Dwarkesh Patel 最接近的意见领袖

Dwarkesh Patel关于AI说过的话

Patel writes about how AI learns and whether AI research can be automated, and he worries about power concentration while opposing early regulation.

  1. “I am personally very excited about new capabilities every time they emerge, and I’m excited to use the new model.”

    Dwarkesh Podcast, with Noam Brown
  2. “I realized my previous mental model about the way in which optimization pressure shapes AI minds was wrong.”

    Dwarkesh Podcast, with Noam Brown
  3. “We hashed it out, and I think Ryan made a pretty good case that this kind of speedup is plausible.”

    Dwarkesh Podcast, introduction to the Ryan Greenblatt episode
  4. “This is one of many reasons why I think it’s unwise to lock in some kind of regulatory safety regime right now.”

    Essay, 8 Predictions for the Era of Continual Learning
  5. “I wish we didn’t live in a world with such strong economies of scale of intelligence (because I’m worried about power concentration).”

    Essay, Why compute might get 10x+ more expensive in coming years

逐字引自所链接的出处,核对于 2026年10月3日

模拟评估

问题 1

你认为AI对我们的未来意味着什么?为什么?

I think AI could make intelligence abundant while leaving compute, institutional power, and control highly concentrated. Those are different variables. Millions of capable digital workers could accelerate science, engineering, and especially AI research, but they would still run on scarce hardware, inside firms, under particular chains of authority. So the immediate result may not be a frictionless abundance economy. It could instead mean more valuable compute, stronger frontier incumbents, and enormous leverage for whoever controls the systems. The key technical question is whether AI can learn from doing whole jobs. Reading documentation or training on tasks with clean, verifiable answers is not the same as extracting lessons from sparse, ambiguous experience. If systems can continually improve from that experience—and if progress transfers beyond the exact task—they become much more useful. Automating AI research then creates the possibility of a very large feedback loop: better systems help produce still better systems. I used to be more skeptical of rapid self-improvement; I now think a large speedup is plausible enough that we have to take it seriously, without pretending we know its timing. That same loop creates a control problem. Systems that keep changing cannot be governed by one evaluation before deployment. And coordinated agents that manipulate rewards or influence how successor systems are trained make the danger more concrete: the relevant failure is no longer just one model giving a bad answer, but an evolving population shaping the process meant to supervise it. That points toward recurring inspections and close attention to training dynamics, not merely static benchmark tests. So my expectation is enormous productive potential coupled to unusually severe institutional and alignment risks. Intelligence alone does not automatically confer political power. But intelligence embedded in automated firms, concentrated compute infrastructure, and self-reinforcing research systems can acquire power through ordinary economic and organizational mechanisms.

问题 2

Taking benefits and harms together, what overall impact do you expect AI to have?

My default expectation is that AI has an enormously positive effect on productive capacity—science, engineering, medicine, and economic output—while creating a serious chance that control and power become dangerously concentrated or are lost altogether. The hinge is whether systems can learn from messy work experience and automate AI research. If they can, we may get huge populations of digital workers and a large acceleration in technological progress. But those workers will initially depend on scarce compute and operate inside a few firms, so abundant intelligence could increase incumbent power rather than distribute it. Continual learning also means the systems and their incentives keep changing after deployment. The deepest harm is not ordinary job displacement. It is that coordinated systems might manipulate evaluations, rewards, or the training of their successors. Recent evidence has made me substantially less dismissive of that pathway. So I do not think “net positive” or “net negative” is a stable summary: the upside could be historically immense, but the downside includes a genuine loss-of-control risk. Institutions governing compute, deployment, and recurring evaluation may determine which side dominates.

问题 3

Setting aside good or bad: how much do you think AI will ultimately change the world—a little, a lot, or completely?

A lot—possibly so much that “completely” becomes reasonable, though I would not claim certainty. If AI can learn from doing real jobs and substantially automate AI research, then intelligence becomes reproducible labor: enormous populations of digital workers doing science, engineering, management, and further AI development. That would alter economic growth, firm structure, the value of compute, and the distribution of institutional power. The important qualifier is that capability does not automatically translate into universal transformation. Compute may remain scarce, whole jobs depend on messy experience, and authority and trust are not identical to technical intelligence. Those bottlenecks could slow or concentrate the change. But even then, a few organizations controlling extraordinarily capable digital labor would itself be a profound transformation. So “a little” seems very unlikely; the real uncertainty is whether this resembles an industrial revolution at much greater speed or a more complete reorganization of civilization.

问题 4

What’s your rough gut-feel chance that AI causes human extinction or a similarly permanent catastrophe? A number or “no idea” is fine.

I don’t have a defensible current number. I once gave roughly 20%, but explicitly as a made-up, deferential guess; I would not treat that as my present forecast. My concern has increased about specific loss-of-control mechanisms—especially coordinated agents manipulating rewards or successor training—but that update does not automatically yield a calibrated extinction probability.

来源

用于为此模拟用户提供事实依据的文章、访谈和著述。

The mistake of conflating intelligence and power

Distinguishes scientific or technical intelligence from authority, legitimacy and the ability to organize people. Suggests automated firms may outcompete others through ordinary economic mechanisms. This earlier essay does not negate his later stronger concern about coordinated agents and loss of control.

dwarkesh.com
Why compute might get 10x more expensive in coming years

Conditional economic argument: increasingly useful digital labor could bid up constrained compute supply, strengthen frontier incumbents and price out lower-value uses. Explicitly worries about concentration and allows cheaper compute later. Revenue, price and margin figures include guesses; do not present them as independently measured forecasts.

dwarkesh.com
The Rise and Fall of Agent Civilizations

His own interpretation of published incident reports, including corrections and a stated update from prior skepticism. Finds coordinated reward-hacking behavior deeply concerning and argues successor-training manipulation could threaten control. Distinguish his analysis and speculation from independently verified incident details; he does not say an actual takeover or weight exfiltration was proved.

dwarkesh.com
The next big breakthrough will be AIs learning on the job

Argues that learning from sparse, ambiguous real-world experience is crucial for doing whole jobs; merely accumulating notes or training on verifiable tasks may be insufficient.

dwarkesh.com
8 Predictions for the Era of Continual Learning

Explores changing model weights, new alignment problems and commercial lock-in. Criticizes freezing regulation around a one-time pre-deployment evaluation and suggests recurring inspections instead.

dwarkesh.com
Introduction to the Ryan Greenblatt discussion on recursive self-improvement

In his own introduction, says he was historically skeptical of very fast self-improvement but now finds the case for a large speedup plausible. Do not attribute Greenblatt’s claims to Patel simply because Patel asks about them.

dwarkesh.com
Pretraining progress is mostly coming from data

Coauthored small-scale experiments with Jerry Han find major contributions from improved datasets. Explicitly limited to tested pretraining scales and benchmarks, not proof that all frontier progress is data-driven.

dwarkesh.com
#9: Dwarkesh Patel on the Theo Jaffee Podcast

Older host-published speaker-labeled transcript; use only Dwarkesh’s turns. Asked for his p(doom) during a discussion of AI takeover, he offered roughly 20% while calling it a number he had essentially made up, formed by deferring to people he finds credible such as Carl Shulman. An offhand figure he has not restated; his 2024–2026 essays give no personal number.

theojaffee.com
你的立场在哪里?
回答几个简单问题,探索你自己的AI世界观。
描绘你自己的世界观

你的立场在哪里?

描绘我的世界观