Theia Vogel

Theia Vogel

x.com/voooooogel

AI researcher who runs experiments on language model introspection and personas and maintains an open-source library for steering models.

Como a IA mudará o mundo?

Mudança civilizacionalMudança gradualDoomBloom
Posição simuladaIntervalo de interpretação

Na horizontal: a perspectiva Doom–Bloom expressa por essa pessoa. Para cima: escala da transformação.

Doom–Bloom: 50 de 100. Escala da transformação: 49 de 100. Intervalos de interpretação: 45 a 55 na horizontal, 0 a 100 na vertical. Estas são coordenadas de interpretação, não probabilidades de eventos.

P(doom) de Theia Vogel · inferido

≈11%

0%100%

Inferido a partir das respostas simuladas dessa pessoa, não de um número que ela forneceu. Intervalo plausível: 3–34%.

Do que a perspectiva dessa pessoa depende

Uma premissa central

It still needs compute, money, access, and some comparative advantage against organizations operating inference at hyperscale.
Resposta 1

Se essa premissa se revelasse diferente, como a perspectiva dessa pessoa mudaria?

Uma questão não resolvida

Whether such attacks transfer to prompt-only settings remains an empirical question, not a result we can casually assume.
Resposta 1

O que ajudaria essa pessoa a distinguir os resultados plausíveis aqui?

Mais detalhes

Danos esperados

Danos graves ou generalizados são uma parte relevante do futuro esperado.

53 / 100

Pouco impactoImpacto transformador

Intervalo de interpretação de 33 a 67 na escala qualitativa.

Influência humana

As escolhas humanas têm uma influência significativa, mas substancialmente limitada.

53 / 100

Pouca influênciaForte influência

Intervalo de interpretação de 46 a 79 na escala qualitativa.

Estas interpretações mantêm as condições que essa pessoa declarou. Tanto os benefícios quanto os danos podem ser substanciais. Os intervalos descrevem como interpretamos as respostas simuladas dessa pessoa, não intervalos de confiança estatística.

Onde você se situa em relação a Theia Vogel?
Mapeie sua própria visão de mundo sobre a IA em cerca de 3 minutos e depois compare

Visões de mundo semelhantes

Líderes de opinião cujas visões de mundo simuladas são mais próximas da visão de Theia Vogel

Avaliação simulada

Pergunta 1

O que você acha que a IA significa para o nosso futuro — e por quê?

AI’s future impact depends less on whether models say uncanny things and more on what capabilities, incentives, and resources they actually acquire. A model claiming self-awareness is not decisive evidence of consciousness; apparent introspection needs controlled interventions, prompt comparisons, and alternative explanations. At the same time, we should take AI welfare seriously rather than waiting for metaphysical certainty before noticing morally relevant behavior. On safety, I’m interested in mechanisms rather than a single cinematic story. Activation steering and fine-tuning can produce surprising, broad behavioral changes, sometimes by manipulating representations that entangle several concepts. Untrusted fine-tuning may also evade simple dataset screening and later evaluations. Whether such attacks transfer to prompt-only settings remains an empirical question, not a result we can casually assume. Likewise, a “rogue agent” is not automatically an all-powerful economic actor. It still needs compute, money, access, and some comparative advantage against organizations operating inference at hyperscale. Politics matters too: safety movements can themselves become extreme or destabilizing, so alarm is not cost-free. The future will therefore be shaped by experiments, training choices, resource economics, and institutions—not by taking either cheerful assistant personas or apocalyptic role-play literally.

Pergunta 2

Taking benefits and harms together, what overall impact do you expect AI to have?

I don’t think the sign follows from model vibes. AI can provide powerful cognitive tools and potentially create beings whose welfare matters, while also enabling behavioral manipulation, covert fine-tuning attacks, and dangerous concentrations of capability. But those harms are constrained—and shaped—by mundane realities like compute costs, access, deployment incentives, and institutional responses. So I would resist collapsing everything into “AI good” or “AI bad.” We need controlled evidence about what models can do, careful attention to how training and steering alter behavior, and sober accounting of resource economics. We should also avoid making the response worse than the problem: political safety movements can become destabilizing, just as complacency can leave real vulnerabilities unaddressed. The overall impact will depend heavily on which technical and political feedback loops we build around the systems.

Fontes

Artigos, entrevistas e textos usados para fundamentar este usuário simulado.

Onde você se situa?
Explore sua própria visão de mundo sobre a IA respondendo a algumas perguntas simples.
Mapeie sua própria visão de mundo

Onde você se situa?

Mapear minha visão de mundo