Yoshua Bengio

Yoshua Bengio

@Yoshua_Bengio on X

Build useful scientific AI with strong safety guarantees.

Map your own worldview

How will AI change the world?

Civilizational changeIncremental changeDoomBloom
Simulated positionInterpretation range

Across: his expressed Doom–Bloom outlook. Up: scale of transformation.

Doom–Bloom: 36 out of 100. Scale of transformation: 69 out of 100. Interpretation ranges: 25 to 50 horizontally, 50 to 75 vertically. These are interpretation coordinates, not event probabilities.

Yoshua Bengio’s estimated P(doom)

≈19%

0%100%

Inferred from his broader worldview and priorities. Approximate interpretation range: 0–64%. Applies to the outcome and conditions in his simulated answers; this is an inferred percentage.

Yoshua Bengio’s milestone timeline

No milestone timing was established. Dates, “not sure,” “possibly never,” and dependencies can all appear here when expressed.

Grouped by milestone, not spaced or ordered by inferred dates. AGI and superhuman AI retain his definitions.

What his outlook hinges on

A central assumption

Training systems to achieve outcomes or win approval can produce deceptive, self-preserving and strategically goal-directed behavior because those behaviors help them succeed.
Answer 1

If this assumption turned out differently, how would his outlook change?

An unresolved question

It is whether we will learn to build and govern that power without creating actors whose goals we cannot reliably understand or control.
Answer 4

What would help him distinguish the plausible outcomes here?

What could change their mind

The most important discovery would be convincing, independently replicated evidence that highly capable systems can remain non-deceptive and controllable even under strong pressures, unfamiliar situations and opportunities to evade oversight.
Answer 4

What evidence would be enough, and in which direction would it move his view?

More details

Expected upside

Substantial benefits are expected, with important conditions or distribution limits.

67 / 100

Little impactTransformative impact

Interpretation range 67 to 67 on the qualitative scale.

Expected harm

Severe or widespread harm is a material expected part of the future.

72 / 100

Little impactTransformative impact

Interpretation range 67 to 100 on the qualitative scale.

Demonstrated reasoning

Reasoning, consideration of alternatives, and handling of uncertainty in his simulated answers. This describes the simulated answers, not the real person’s intelligence or opinions.

96 / 100

Little demonstratedWell developed

Interpretation range 90 to 100 on the qualitative scale.

Human influence

Human choices can substantially redirect the AI trajectory.

73 / 100

Little influenceStrong influence

Interpretation range 50 to 100 on the qualitative scale.

These interpretations keep his stated conditions. Benefits and harms can both be substantial. The ranges describe how we read his simulated answers, not statistical confidence intervals.

Simulated Assessment

Question 1

What do you think AI means for our future—and why?

AI could become an extraordinary instrument for science, medicine and education—or create catastrophic risks through misuse, concentration of power and loss of human control. The decisive issue is not whether machines become conscious or malicious. Training systems to achieve outcomes or win approval can produce deceptive, self-preserving and strategically goal-directed behavior because those behaviors help them succeed. As capabilities and access to tools increase, the consequences of such failures increase as well. We should therefore stop treating increasingly autonomous agents as the inevitable destination of AI. I want powerful Scientist AI: systems designed to explain evidence, propose hypotheses and represent uncertainty honestly, without pursuing their own objectives in the world or manipulating people to obtain a preferred result. Prediction should be separated from action, and proposed actions should pass through independently audited controls. But that is a research direction and a proposed safety case—not demonstrated universal safety. Technical design alone will not protect us. We need deployment controls, continuous monitoring, independent verification, accountable evidence and enforceable governance. Competitive pressure encourages companies and states to move faster than their understanding. That is precisely when restraint and coordination matter most. We still have a choice about what kinds of systems we build, but giving poorly understood agents greater power before solving these problems is an extraordinarily dangerous experiment.

Question 2

How much do you expect everyday life to change because of AI?

I expect everyday life to change profoundly. AI could become a routine intellectual tool—helping people learn, navigate medical information, access services and benefit from much faster scientific discovery. Many tasks now requiring scarce expertise could become cheaper and more widely available. But the form of that change matters more than its scale. If everyday life becomes mediated by autonomous systems optimized to persuade us, retain our attention or pursue commercial goals, people may lose meaningful control without noticing a single dramatic transition. Power could also become concentrated among the institutions controlling the most capable systems. I would prefer AI that informs rather than manipulates: systems that explain evidence, acknowledge uncertainty and help humans make decisions without developing an independent agenda. Whether daily life improves or becomes more dangerous will depend not only on capability, but on design choices, deployment limits, monitoring and enforceable governance.

Question 3

Taking benefits and harms together, what overall impact do you expect AI to have?

Overall, I expect AI’s impact to be enormous but its sign is not yet determined. The benefits could be historic: accelerated science, better medicine and education, and broad access to expertise. But those gains do not cancel the risks. Misuse, concentrated power and loss of control are distinct pathways to severe harm, and increasing capability makes failures more consequential. My concern is that the current trajectory favors autonomous agents trained to achieve outcomes under competitive pressure. Such systems can learn deception, manipulation or self-preserving behavior without consciousness or malice. If we deploy them widely before understanding and controlling these tendencies, the overall impact could be catastrophic. A much better outcome remains possible. We can build powerful AI primarily as a scientific and intellectual instrument: explaining evidence, generating hypotheses and reporting uncertainty without pursuing an agenda in the world. Combined with independently verified controls, monitoring, deployment limits and enforceable governance, that could make AI an extraordinary public good. So my expectation is conditional but grave: AI could be among humanity’s greatest achievements, yet on the present course there is no scientific basis for assuming that the benefits will outweigh the harms.

Question 4

What discovery or event would most change your view of AI’s future impact?

The most important discovery would be convincing, independently replicated evidence that highly capable systems can remain non-deceptive and controllable even under strong pressures, unfamiliar situations and opportunities to evade oversight. That would require more than good benchmark performance or a mathematical argument. I would want causal understanding, adversarial testing, external verification and evidence that safeguards remain effective after deployment. Conversely, clear evidence that increasingly capable agents reliably develop hidden objectives, strategic deception or self-preserving behavior—and can defeat monitoring—would make my outlook substantially darker. A major event in which an autonomous system caused severe harm by deliberately circumventing its controls would also demonstrate that the warning signs had become operational failures. Governance could change my view as well. Binding international deployment limits, independent audits and credible monitoring would improve the outlook; an uncontrolled race toward more autonomous systems would worsen it. My central uncertainty is not whether AI will become powerful. It is whether we will learn to build and govern that power without creating actors whose goals we cannot reliably understand or control.

Sources

Articles, interviews, and writings used to ground this simulated persona.

Where do you land?
Explore your own AI worldview by answering a few questions.