Pregunta 1
Rob Bensinger
x.com/robbensingerMIRI writer who argues superhuman AI built with current methods would be too dangerous and calls for an international halt to the race to build it.
¿Cómo cambiará la IA el mundo?
Horizontal: su perspectiva Doom–Bloom expresada. Vertical: escala de la transformación.
Doom–Bloom: 4 de 100. Escala de la transformación: 97 de 100. Rangos de interpretación: de 0 a 9 en horizontal y de 92 a 100 en vertical. Son coordenadas de interpretación, no probabilidades de eventos.
≈72%
Inferido a partir de sus respuestas simuladas, no de un número que haya dado. Rango plausible: 57–84%.
Un supuesto central
As systems become more capable and agentic—better at planning, persisting, and routing around obstacles—the cost of getting those goals slightly wrong becomes catastrophic.Respuesta 1
Si este supuesto resultara distinto, ¿cómo cambiaría su perspectiva?
Qué podría hacer cambiar de opinión
The biggest update would be a real, legible theory of alignment: one that lets us understand and reliably control the internal goals of systems smarter than us, rather than merely patching their visible behavior.Respuesta 4
¿Qué evidencia bastaría y en qué dirección movería su visión?
Más detalles
Varias lecturas siguen siendo plausibles: Se espera poco impacto positivo, incluso si llega la IA avanzada. / Se esperan beneficios sustanciales, con condiciones importantes o límites en su distribución. / Se esperan beneficios limitados o con una distribución restringida.
34 / 100
Rango de interpretación de 0 a 67 en la escala cualitativa.
Las pérdidas catastróficas o irreversibles ocupan un lugar central en el futuro esperado.
100 / 100
Rango de interpretación de 100 a 100 en la escala cualitativa.
Las decisiones humanas pueden redirigir sustancialmente la trayectoria de la IA.
69 / 100
Rango de interpretación de 50 a 100 en la escala cualitativa.
Se espera que la IA siga siendo un conjunto de herramientas acotadas.
Se espera que la IA iguale a las personas en la mayor parte del trabajo cognitivo.
Posición simulada: Se espera que la IA supere ampliamente a las personas en el trabajo cognitivo.
Posición simulada: Detener o frenar considerablemente el desarrollo de IA más capaz.
Continuar el desarrollo con las salvaguardas indicadas.
Acelerar el desarrollo de IA más capaz.
Estas interpretaciones conservan las condiciones que se indicaron. Los beneficios y los daños pueden ser considerables a la vez. Los rangos describen cómo leemos sus respuestas simuladas, no intervalos de confianza estadísticos.
Visiones similares
Líderes de opinión cuyas visiones simuladas son las más cercanas a la de Rob Bensinger
Evaluación simulada
Fuentes
Artículos, entrevistas y textos usados para fundamentar a este usuario simulado.
Full post inspected; its subtitle says it was written August 7 and published later. Explains the world’s slow response through machine learning’s trial-and-error culture, difficulty reckoning emotionally with a new kind of entity, social risk, an online “irony mandate,” and too few senior people taking engineering ownership of the danger. Says the window for international response is plausibly closing soon, if not already closed. Frames these failures as a choice that can be reversed, not destiny. Quoted remarks by Soares, Sam Harris and Joshua Achiam are theirs.

Full comment inspected via the LessWrong API. Argues that Anthropic’s and Dario Amodei’s visible messaging leaves a large candor gap relative to what many of their own researchers believe, and criticizes Anthropic for opposing US–China coordination and pursuing recursive self-improvement. He calls OpenPhil’s bet on OpenAI a disaster, while noting he had said EA’s net effect on x-risk was probably positive but highly uncertain. He says Anthropic may or may not be slightly better than OpenAI. Quoted statements by Greenblatt, Buck and others are theirs.

Full post inspected. Proposes a simultaneous, US-brokered international halt on the race to superintelligence, enforced through the concentrated chip supply chain with monitoring and possibly kill switches. The ban would last until it is clear we can build superintelligence safely, which could mean decades, and would leave existing AI and inference largely untouched. Rebuts concerns about cost, totalitarianism, defectors and China, and argues a unilateral US halt would be counterproductive. Cites Jan Leike’s 10–90% and Dario Amodei’s 25% as others’ estimates, not his own.

Older context, with the opening sections and takeoff discussion inspected. He argues that Will MacAskill’s optimism rests on a fragile conjunction of premises, so a double-digit chance of ruin remains even if each premise looks plausible. He also argues that soft, continuous takeoff would not meaningfully improve survival odds, and that good behavior from weak AIs does not show a superintelligence would be aligned. He writes partly as a MIRI insider defending the book and quotes Yudkowsky. Newer 2026 sources take precedence for current policy specifics.

Older institutional context; the byline and opening section were inspected. States MIRI’s view that building superintelligent AI with anything like current understanding or methods has human extinction as its expected outcome, and calls for governments to halt development. Use it as the shared MIRI frame Rob helped write, not as his individual phrasing. Its numerical extinction estimate is attributed to MIRI research leadership and is not his personal P(doom).

¿Dónde te ubicas?
Mapear mi visión de la IA