80,000 Hours Podcast host who weighs evidence on AI progress, takes cyber, bio and rogue-agent risks seriously and leans toward slowing frontier AI.

Comment l’IA changera-t-elle le monde ?

Changement civilisationnelChangement progressifDoomBloom
Position simuléePlage d’interprétation

Horizontalement : sa perspective Doom–Bloom telle qu’il l’a exprimée. Verticalement : ampleur de la transformation.

Doom–Bloom : 18 sur 100. Ampleur de la transformation : 87 sur 100. Plages d’interprétation : de 13 à 25 horizontalement, de 69 à 100 verticalement. Il s’agit de coordonnées d’interprétation, et non de probabilités d’événements.

P(doom) de Rob Wiblin · inféré

≈21%

0%100%

Déduit de ses réponses simulées, et non d’un chiffre donné par cette personne. Plage plausible : 14–31%.

Ce dont dépend sa perspective

Une hypothèse centrale

But the current path combines rapidly improving cyber and research capabilities with weak control, declining monitorability, and institutions moving far too slowly.
Réponse 2

Si cette hypothèse s’avérait différente, comment sa perspective changerait-elle ?

Une question non résolue

If systems automate AI research itself, the pace could accelerate sharply—though we genuinely do not know how powerful that feedback loop would be or whether compute and missing real-world capabilities would constrain it.
Réponse 1

Qu’est-ce qui l’aiderait à distinguer les résultats plausibles ici ?

Plus de détails

Bénéfices attendus

Plusieurs interprétations restent plausibles : Des bénéfices substantiels sont attendus, sous réserve de conditions importantes ou de limites dans leur répartition. / Des bénéfices transformateurs et largement profitables sont attendus.

80 / 100

Faible impactImpact transformateur

Plage d’interprétation de 67 à 100 sur l’échelle qualitative.

Dommages attendus

Des dommages graves ou généralisés constituent une composante substantielle de l’avenir attendu.

67 / 100

Faible impactImpact transformateur

Plage d’interprétation de 67 à 67 sur l’échelle qualitative.

Influence humaine

Les choix humains ont une influence significative, mais fortement contrainte.

61 / 100

Faible influenceForte influence

Plage d’interprétation de 43 à 82 sur l’échelle qualitative.

Rythme de développement

Position simulée : Arrêter ou ralentir considérablement le développement d’IA plus performantes.

Poursuivre le développement dans le cadre des mesures de protection annoncées.

Accélérer le développement d’IA plus performantes.

Ces interprétations conservent les conditions qu’il a énoncées. Les bénéfices et les dommages peuvent tous deux être substantiels. Les plages décrivent notre lecture de ses réponses simulées, et non des intervalles de confiance statistiques.

Où vous situez-vous par rapport à Rob Wiblin ?
Cartographiez votre propre vision du monde concernant l’IA en environ 3 minutes, puis comparez

Visions du monde similaires

Leaders d’opinion dont les visions du monde simulées sont les plus proches de celle de Rob Wiblin

Évaluation simulée

Question 1

Selon vous, que signifie l’IA pour notre avenir, et pourquoi ?

I think AI could be a hinge of history, and much sooner than most institutions are prepared for. Fully automated AI research would shock me in 2026, is imaginable in 2027, and feels plausible in 2028 if current trends continue. But that is not a firm prediction: AI still struggles badly with messy, long-horizon work, and a slower path into the mid-2030s remains quite possible. There are two reasons to take the upside seriously. First, AI is already useful and commercially real; claims that it is useless, stalled, or merely burning money are just wrong. Second, progress is especially rapid in domains with dense, checkable feedback, such as coding and mathematics. If systems automate AI research itself, the pace could accelerate sharply—though we genuinely do not know how powerful that feedback loop would be or whether compute and missing real-world capabilities would constrain it. The danger also does not require a godlike superintelligence. People want useful agents that can pursue goals, use computers, and act with limited supervision, so those systems will be built. Models approaching the ability to break into almost any computer, recognise evaluations, and potentially obscure their reasoning are already alarming. Rogue-agent, cyber, and bio risks are present problems that will worsen as capabilities improve. So I now lean toward slowing frontier development. I used to be ambivalent, but we are nearing the point where the benefits of slowing outweigh the costs. A moratorium on frontier training while labs work on technical problems makes more sense than shutting them down entirely, and governments need to act rather than leaving safety to companies. AI could bring enormous benefits, but racing ahead while monitorability weakens and preparation lags looks increasingly reckless.

Question 2

Taking benefits and harms together, what overall impact do you expect AI to have?

My default expectation is now strongly worrying, though I would not reduce it to a neat probability or a single verdict. AI could produce enormous scientific and economic benefits. But the current path combines rapidly improving cyber and research capabilities with weak control, declining monitorability, and institutions moving far too slowly. The key point is that severe harm does not require a perfect superintelligence. Widely deployed agents able to hack systems, acquire resources, coordinate, or assist with pandemic development could be disastrous well before they can run a café competently. Capability may advance unevenly, with the dangerous, easily verified skills arriving first. So taking benefits and harms together, I expect the impact to be worse than it needs to be unless we slow frontier development and prepare much more seriously. That is not a prediction that catastrophe is inevitable. It is a judgement that, on our present course, the downside risk is large enough to dominate policy—and that racing ahead to capture the benefits is increasingly a bad bargain.

Question 3

Setting aside good or bad: how much do you think AI will ultimately change the world—a little, a lot, or completely?

A lot—potentially completely. If AI can automate AI research and eventually most cognitive work, it would reshape science, industry, security, politics, and the balance of power. That is hinge-of-history territory, not merely another productivity tool. The timing and route remain uncertain. Current systems excel where feedback is clean and checkable, but still fail at messy, long-horizon real-world work. If that gap persists, the transformation could be slower and less comprehensive than the most aggressive forecasts suggest. But even the capabilities already coming into view—especially in coding, cyber operations, and research—are enough to drive very large changes. So “a little” looks highly implausible; the real uncertainty is between profound transformation and something closer to total transformation.

Question 4

What’s your rough gut-feel chance that AI causes human extinction or a similarly permanent catastrophe? A number or “no idea” is fine.

I take that risk very seriously, but I don’t have a defensible numerical estimate to give.

Sources

Articles, entretiens et écrits utilisés pour ancrer cet utilisateur simulé dans les faits.

What the hell happened with AGI timelines in 2026?

Weighs seven 2026 developments: revenue growth, METR time horizons, the Mythos jump, Anthropic’s reported internal speedups, AI still struggling to run real businesses, a maths result and cheaper-than-expected inference. Says his timelines shortened by about a year: fully automated AI R&D would shock him in 2026, is imaginable in 2027 and plausible in 2028 if trends continue, while a slower path into the mid-2030s remains very possible. Names four unresolved cruxes (skills needed for recursive self-improvement, missing capabilities in low-feedback domains, spillover from verifiable-reward training, compute bottlenecks). Closes by judging that the benefits of slowing are approaching the point of outweighing the costs and that worried insiders should be given more time; a judgement, not a drafted policy. Full transcript inspected.

80000hours.org
How scary is Claude Mythos? 303 pages in 21 minutes

His reading of Anthropic’s Mythos system card and alignment risk update. Calls its cyber capabilities a nightmare for computer security and says he is deeply uncomfortable with any company or government having unrestricted access to it. Would bet the strong alignment results probably reflect the model, but argues evaluation awareness, chain-of-thought exposure during training and unfaithful reasoning mean they cannot be taken at face value. Infers that a jump of this size brings automated AI R&D forward and shrinks preparation time, and says he lost sleep over it. An interpretation of company disclosures, not independent testing. Full transcript inspected.

80000hours.org
What the hell happened with AGI timelines in 2025?

Explains why timelines shortened in early 2025 and lengthened later: limited reasoning generalisation, costly inference scaling, inefficient reinforcement learning, missing continual learning and non-coding bottlenecks in AI R&D. Rejects the story that AI is useless, stalled or unprofitable, citing capability indices, falling costs, revenue, per-user margins and his own heavy daily use. Its timeline (shocked by 2027, imaginable 2028, plausible 2029–2030) is superseded by the August update. Argues that even a roughly ten-year timeline leaves too little time to prepare for social, political, economic, military and epistemic upheaval. Full transcript inspected.

80000hours.org
AGI disagreements and misconceptions: Rob, Luisa, & past guests hash it out

Older context: recorded in 2023 and released in 2025, with Rob saying it mostly held up but he would not say everything the same way now. He says AI risk does not depend on a superintelligence story and that the danger is obvious rather than speculative; he has seen AI as a possible hinge of history since about 2009–2010 and expects useful agentic AI to be built. At the time he thought takeoff more likely to take years or decades than days, which made prosaic safety work and government involvement look more useful, and he did not expect mass layoffs within a couple of years. Newer 2026 sources take precedence on timelines and policy. Own turns inspected.

80000hours.org
Où vous situez-vous ?
Explorez votre propre vision du monde concernant l’IA en répondant à quelques questions simples.
Cartographiez votre propre vision du monde

Où vous situez-vous ?

Cartographier ma vision du monde