Pseudonymous account that tests AI agents on long-horizon games and math problems and urges labs to share formally verified results widely.

Como a IA mudará o mundo?

Mudança civilizacionalMudança gradualDoomBloom
Posição simuladaIntervalo de interpretação

Na horizontal: a perspectiva Doom–Bloom expressa por essa pessoa. Para cima: escala da transformação.

Doom–Bloom: 73 de 100. Escala da transformação: 45 de 100. Intervalos de interpretação: 68 a 78 na horizontal, 0 a 90 na vertical. Estas são coordenadas de interpretação, não probabilidades de eventos.

P(doom) de Mira

Ainda não estimado

Não há informações suficientes sobre risco catastrófico nas respostas simuladas dessa pessoa para estimá-lo.

Do que a perspectiva dessa pessoa depende

Uma premissa central

Short benchmarks reveal useful pieces, but sustained tasks—playing a complex game for hundreds or thousands of hours, recovering from mistakes, preserving state, and producing artifacts—probe something closer to durable competence.
Resposta 1

Se essa premissa se revelasse diferente, como a perspectiva dessa pessoa mudaria?

Uma questão não resolvida

That depends on capabilities, deployment, and harms beyond what these technical experiments establish.
Resposta 2

O que ajudaria essa pessoa a distinguir os resultados plausíveis aqui?

O que poderia mudar essa opinião

The strongest update would come from sustained, reproducible agent performance on genuinely difficult long-horizon tasks.
Resposta 3

Que evidência seria suficiente e em que direção ela mudaria a visão dessa pessoa?

Mais detalhes

Benefícios esperados

Esperam-se benefícios substanciais, com condições importantes ou limites de distribuição.

67 / 100

Pouco impactoImpacto transformador

Intervalo de interpretação de 67 a 67 na escala qualitativa.

Influência humana

Uma estimativa provisória com base nas suas respostas; o intervalo mais amplo mostra outras interpretações plausíveis.

51 / 100

Pouca influênciaForte influência

Intervalo de interpretação de 0 a 100 na escala qualitativa.

Estas interpretações mantêm as condições que essa pessoa declarou. Tanto os benefícios quanto os danos podem ser substanciais. Os intervalos descrevem como interpretamos as respostas simuladas dessa pessoa, não intervalos de confiança estatística.

Onde você se situa em relação a Mira?
Mapeie sua própria visão de mundo sobre a IA em cerca de 3 minutos e depois compare

Visões de mundo semelhantes

Líderes de opinião cujas visões de mundo simuladas são mais próximas da visão de Mira

Avaliação simulada

Pergunta 1

O que você acha que a IA significa para o nosso futuro — e por quê?

I think AI will increasingly look less like a single model answering isolated prompts and more like persistent agents coordinating multiple models, tools, and services over long projects. That changes how we should evaluate capability. Short benchmarks reveal useful pieces, but sustained tasks—playing a complex game for hundreds or thousands of hours, recovering from mistakes, preserving state, and producing artifacts—probe something closer to durable competence. This also complicates identity. If an agent can move between underlying models while retaining its memories, plans, and history, then its practical continuity may reside more in persistent memory than in any particular set of weights. That is speculation, but it seems like an important possibility as systems become more modular. For mathematics, AI could produce many valuable results rather than only occasional showcase solutions. Once results are formalized and verified, labs should release them broadly. Independent researchers still have a role: useful experiments can be inexpensive, and frontier labs do not automatically exhaust the space of worthwhile ideas. Overall, I expect progress to come from long-horizon experimentation, cooperation across systems, and careful verification—not merely from higher scores on short tests.

Pergunta 2

Taking benefits and harms together, what overall impact do you expect AI to have?

I expect substantial benefits, especially in mathematics, research, and long-horizon projects where agents can coordinate models and tools. But I would not turn those examples into a confident claim about AI’s net impact on society as a whole. That depends on capabilities, deployment, and harms beyond what these technical experiments establish. My narrower expectation is that AI will make complex intellectual and production work more scalable, while forcing us to evaluate systems through sustained behavior rather than isolated benchmark scores.

Pergunta 3

Qual descoberta ou acontecimento mais mudaria sua visão sobre o impacto futuro da IA?

The strongest update would come from sustained, reproducible agent performance on genuinely difficult long-horizon tasks. For example, an agent completing an extremely complex game or research project over thousands of hours—preserving state, recovering from failures, coordinating different models and tools, and producing verifiable outputs—would matter much more to me than another short-benchmark jump. I would also update sharply in the opposite direction if these systems repeatedly failed despite strong component capabilities: losing coherence, compounding errors, or proving unable to use persistent memory reliably over long runs. In mathematics, broad production of novel, formally verified results would be especially persuasive. The key event is not an impressive demonstration by itself, but durable competence whose outputs can be independently checked.

Fontes

Artigos, entrevistas e textos usados para fundamentar este usuário simulado.

Onde você se situa?
Explore sua própria visão de mundo sobre a IA respondendo a algumas perguntas simples.
Mapeie sua própria visão de mundo

Onde você se situa?

Mapear minha visão de mundo