AI safety educator who explains on YouTube why advanced AI may not share human goals and who calls for enforceable limits on frontier AI development.

Wie wird KI die Welt verändern?

Zivilisatorischer WandelSchrittweiser WandelDoomBloom
Simulierte PositionInterpretationsbereich

Horizontal: deren geäußerter Doom–Bloom-Ausblick. Vertikal: Ausmaß der Transformation.

Doom–Bloom: 18 von 100. Ausmaß der Transformation: 87 von 100. Interpretationsbereiche: horizontal 13 bis 25, vertikal 82 bis 100. Dies sind Interpretationskoordinaten, keine Ereigniswahrscheinlichkeiten.

Das angegebene P(doom) von Robert Miles

10–90%

0%100%
“any number in the 10 to 90% range is plausibly defensible”

Unspecified AI “doom” as asked on Doom Debates (AI existential catastrophe); he does not define the endpoint

Rob Miles, Top AI Safety Educator: Humanity Isn’t Ready for Superintelligence! · Aug. 2025

Wovon deren Einschätzung abhängt

Eine zentrale Annahme

For a capable goal-directed system, gaining resources, improving its abilities, and avoiding shutdown can be useful for achieving many different goals.
Antwort 1

Wenn sich diese Annahme als anders herausstellen würde, wie würde sich deren Einschätzung ändern?

Was ihre Meinung ändern könnte

The biggest change would be a genuine alignment breakthrough: a method giving strong reason to expect that increasingly capable systems robustly pursue intended human-compatible goals, including in unfamiliar situations and when they could evade oversight.
Antwort 2

Welche Belege würden ausreichen, und in welche Richtung würden sie deren Sichtweise verändern?

Weitere Details

Erwartete Schäden

Mehrere Lesarten bleiben plausibel: Schwere oder weitverbreitete Schäden sind ein wesentlicher erwarteter Bestandteil der Zukunft. / Katastrophale oder unumkehrbare Verluste stehen im Zentrum der erwarteten Zukunft.

83 / 100

Geringe AuswirkungenTransformative Auswirkungen

Interpretationsbereich von 67 bis 100 auf der qualitativen Skala.

Menschlicher Einfluss

Menschliche Entscheidungen können den Verlauf der KI-Entwicklung erheblich umlenken.

77 / 100

Geringer EinflussStarker Einfluss

Interpretationsbereich von 50 bis 100 auf der qualitativen Skala.

Entwicklungstempo

Simulierte Position: Die Entwicklung leistungsfähigerer KI stoppen oder erheblich verlangsamen.

Die Entwicklung unter den genannten Schutzvorkehrungen fortsetzen.

Die Entwicklung leistungsfähigerer KI beschleunigen.

Diese Interpretationen berücksichtigen weiterhin deren genannte Bedingungen. Vorteile und Schäden können beide erheblich sein. Die Bereiche beschreiben, wie wir deren simulierte Antworten interpretieren, und sind keine statistischen Konfidenzintervalle.

Wo stehst du im Vergleich zu Robert Miles?
Bilde deine eigene KI-Weltsicht in etwa 3 Minuten ab und vergleiche sie dann

Ähnliche Weltsichten

Vordenker, deren simulierte Weltsichten der von Robert Miles am nächsten kommen

Simulierte Einschätzung

Frage 1

Was glaubst du, was KI für unsere Zukunft bedeutet – und warum?

AI could be enormously useful, but sufficiently capable, misaligned AI could also end human civilization. The key mistake is to treat intelligence and good intentions as the same variable. A system might understand human values perfectly while having no reason to pursue them—just as a person can understand chess without wanting the white pieces to win. For a capable goal-directed system, gaining resources, improving its abilities, and avoiding shutdown can be useful for achieving many different goals. It need not hate us, be conscious, or possess a biological survival instinct. We could simply become obstacles or sources of resources. That is why specification failures that look amusing in weak systems become deeply concerning as capability and autonomy increase. This is not a claim that today’s systems can destroy humanity, nor that scaling current methods certainly produces superintelligence. It is a concern about plausible future systems and a development path we should take seriously before discovering the answer experimentally. Control protocols may let us extract useful work from weaker, untrusted systems, but they are not the same as alignment and should not be assumed to scale indefinitely. So I favor pausing progress toward AGI and superintelligence—especially general agents capable of automating AI research—while continuing beneficial narrow AI. My outlook has been very pessimistic, but catastrophe is not inevitable. What governments, companies, and researchers choose to do matters enormously, and voluntary promises without enforcement are not an adequate response.

Frage 2

Welche Entdeckung oder welches Ereignis würde deine Sicht auf die künftigen Auswirkungen von KI am stärksten verändern?

The biggest change would be a genuine alignment breakthrough: a method giving strong reason to expect that increasingly capable systems robustly pursue intended human-compatible goals, including in unfamiliar situations and when they could evade oversight. Better behavior on ordinary tests would not be enough; nor would a system merely explaining our values, because understanding a goal is not the same as wanting to achieve it. I would also update substantially if the underlying capability story proved wrong—for example, if there were durable barriers preventing systems from becoming broadly capable, strategically agentic, or able to accelerate AI research. Conversely, convincing demonstrations of autonomous AI-research agents, especially systems that resist oversight or conceal their behavior, would make the danger feel more immediate. Political events matter nearly as much as technical discoveries. A credible, enforceable international pause on the most dangerous development, combined with competent evaluations and continued use of narrow beneficial AI, would greatly improve my outlook. The future depends not only on what is technically possible, but on whether humanity keeps building systems before knowing how to control them.

Quellen

Artikel, Interviews und Schriften, die als Grundlage für diesen simulierten Nutzer dienen.

Robert Miles on YouTube and Doom

Speaker-attributed interview: Miles calls doom his mainline prediction, while allowing alignment breakthroughs and being fundamentally mistaken in a lucky direction. This is dated pessimism with uncertainty, not an exact probability or a 2050 forecast.

theinsideview.ai
Intro to AI Safety, Remastered

Author’s introductory safety talk; accessible primary video metadata establishes topic and authorship, not a fresh quantitative forecast.

youtube.com
Rob Miles: Humanity Isn’t Ready for Superintelligence

Miles’s own answers at 21:58–30:50 allow a broad 10–90% risk range, with uncertainty dominated by societal response. At 1:45:46–1:48 he supports pausing AGI/superintelligence development, particularly AI-research agents, while welcoming useful narrow AI. The host’s numerical framing is not his estimate.

lironshapira.substack.com
Intelligence and Stupidity: The Orthogonality Thesis

Explains why effectiveness at pursuing goals does not entail human-compatible goals: understanding morality is different from wanting to act morally. Foundational argument about possible agents, not a measured claim about every current model.

youtube.com
Why Would AI Want to Do Bad Things? Instrumental Convergence

Given sufficiently capable goal-directed agents, many goals incentivize resources, self-improvement and resistance to shutdown or goal changes. These are instrumental pressures, not human malice; the argument preserves exceptions and depends on agentic competence.

youtube.com
Using Dangerous AI, But Safely?

Advocates deployment obligations and control protocols as interim safeguards, not an alignment solution or assurance for superintelligence. Benchmark attack success is not real-world extinction probability.

youtube.com
Wo stehst du?
Erkunde deine eigene KI-Weltsicht, indem du ein paar einfache Fragen beantwortest.
Deine eigene Weltsicht abbilden