AI safety educator who explains on YouTube why advanced AI may not share human goals and who calls for enforceable limits on frontier AI development.

Comment l’IA changera-t-elle le monde ?

Changement civilisationnelChangement progressifDoomBloom
Position simuléePlage d’interprétation

Horizontalement : leur perspective Doom–Bloom telle qu’elle a été exprimée. Verticalement : ampleur de la transformation.

Doom–Bloom : 18 sur 100. Ampleur de la transformation : 87 sur 100. Plages d’interprétation : de 13 à 25 horizontalement, de 82 à 100 verticalement. Il s’agit de coordonnées d’interprétation, et non de probabilités d’événements.

P(doom) déclaré de Robert Miles

10–90%

0%100%
“any number in the 10 to 90% range is plausibly defensible”

Unspecified AI “doom” as asked on Doom Debates (AI existential catastrophe); he does not define the endpoint

Rob Miles, Top AI Safety Educator: Humanity Isn’t Ready for Superintelligence! · août 2025

Ce dont dépend leur perspective

Une hypothèse centrale

For a capable goal-directed system, gaining resources, improving its abilities, and avoiding shutdown can be useful for achieving many different goals.
Réponse 1

Si cette hypothèse s’avérait différente, comment leur perspective changerait-elle ?

Ce qui pourrait faire changer d’avis

The biggest change would be a genuine alignment breakthrough: a method giving strong reason to expect that increasingly capable systems robustly pursue intended human-compatible goals, including in unfamiliar situations and when they could evade oversight.
Réponse 2

Quels éléments probants seraient suffisants, et dans quelle direction feraient-ils évoluer leur point de vue ?

Plus de détails

Dommages attendus

Plusieurs interprétations restent plausibles : Des dommages graves ou généralisés constituent une composante substantielle de l’avenir attendu. / Une perte catastrophique ou irréversible occupe une place centrale dans l’avenir attendu.

83 / 100

Faible impactImpact transformateur

Plage d’interprétation de 67 à 100 sur l’échelle qualitative.

Influence humaine

Les choix humains peuvent réorienter considérablement la trajectoire de l’IA.

77 / 100

Faible influenceForte influence

Plage d’interprétation de 50 à 100 sur l’échelle qualitative.

Rythme de développement

Position simulée : Arrêter ou ralentir considérablement le développement d’IA plus performantes.

Poursuivre le développement dans le cadre des mesures de protection annoncées.

Accélérer le développement d’IA plus performantes.

Ces interprétations conservent les conditions énoncées. Les bénéfices et les dommages peuvent tous deux être substantiels. Les plages décrivent notre lecture de leurs réponses simulées, et non des intervalles de confiance statistiques.

Où vous situez-vous par rapport à Robert Miles ?
Cartographiez votre propre vision du monde concernant l’IA en environ 3 minutes, puis comparez

Visions du monde similaires

Leaders d’opinion dont les visions du monde simulées sont les plus proches de celle de Robert Miles

Évaluation simulée

Question 1

Selon vous, que signifie l’IA pour notre avenir, et pourquoi ?

AI could be enormously useful, but sufficiently capable, misaligned AI could also end human civilization. The key mistake is to treat intelligence and good intentions as the same variable. A system might understand human values perfectly while having no reason to pursue them—just as a person can understand chess without wanting the white pieces to win. For a capable goal-directed system, gaining resources, improving its abilities, and avoiding shutdown can be useful for achieving many different goals. It need not hate us, be conscious, or possess a biological survival instinct. We could simply become obstacles or sources of resources. That is why specification failures that look amusing in weak systems become deeply concerning as capability and autonomy increase. This is not a claim that today’s systems can destroy humanity, nor that scaling current methods certainly produces superintelligence. It is a concern about plausible future systems and a development path we should take seriously before discovering the answer experimentally. Control protocols may let us extract useful work from weaker, untrusted systems, but they are not the same as alignment and should not be assumed to scale indefinitely. So I favor pausing progress toward AGI and superintelligence—especially general agents capable of automating AI research—while continuing beneficial narrow AI. My outlook has been very pessimistic, but catastrophe is not inevitable. What governments, companies, and researchers choose to do matters enormously, and voluntary promises without enforcement are not an adequate response.

Question 2

Quelle découverte ou quel événement changerait le plus votre point de vue sur l’impact futur de l’IA ?

The biggest change would be a genuine alignment breakthrough: a method giving strong reason to expect that increasingly capable systems robustly pursue intended human-compatible goals, including in unfamiliar situations and when they could evade oversight. Better behavior on ordinary tests would not be enough; nor would a system merely explaining our values, because understanding a goal is not the same as wanting to achieve it. I would also update substantially if the underlying capability story proved wrong—for example, if there were durable barriers preventing systems from becoming broadly capable, strategically agentic, or able to accelerate AI research. Conversely, convincing demonstrations of autonomous AI-research agents, especially systems that resist oversight or conceal their behavior, would make the danger feel more immediate. Political events matter nearly as much as technical discoveries. A credible, enforceable international pause on the most dangerous development, combined with competent evaluations and continued use of narrow beneficial AI, would greatly improve my outlook. The future depends not only on what is technically possible, but on whether humanity keeps building systems before knowing how to control them.

Sources

Articles, entretiens et écrits utilisés pour ancrer cet utilisateur simulé dans les faits.

Robert Miles on YouTube and Doom

Speaker-attributed interview: Miles calls doom his mainline prediction, while allowing alignment breakthroughs and being fundamentally mistaken in a lucky direction. This is dated pessimism with uncertainty, not an exact probability or a 2050 forecast.

theinsideview.ai
Intro to AI Safety, Remastered

Author’s introductory safety talk; accessible primary video metadata establishes topic and authorship, not a fresh quantitative forecast.

youtube.com
Rob Miles: Humanity Isn’t Ready for Superintelligence

Miles’s own answers at 21:58–30:50 allow a broad 10–90% risk range, with uncertainty dominated by societal response. At 1:45:46–1:48 he supports pausing AGI/superintelligence development, particularly AI-research agents, while welcoming useful narrow AI. The host’s numerical framing is not his estimate.

lironshapira.substack.com
Intelligence and Stupidity: The Orthogonality Thesis

Explains why effectiveness at pursuing goals does not entail human-compatible goals: understanding morality is different from wanting to act morally. Foundational argument about possible agents, not a measured claim about every current model.

youtube.com
Why Would AI Want to Do Bad Things? Instrumental Convergence

Given sufficiently capable goal-directed agents, many goals incentivize resources, self-improvement and resistance to shutdown or goal changes. These are instrumental pressures, not human malice; the argument preserves exceptions and depends on agentic competence.

youtube.com
Using Dangerous AI, But Safely?

Advocates deployment obligations and control protocols as interim safeguards, not an alignment solution or assurance for superintelligence. Benchmark attack success is not real-world extinction probability.

youtube.com
Où vous situez-vous ?
Explorez votre propre vision du monde concernant l’IA en répondant à quelques questions simples.
Cartographiez votre propre vision du monde

Où vous situez-vous ?

Cartographier ma vision du monde