質問1
Thomas G. Dietterich
x.com/tdietterichOregon State machine learning professor emeritus who works on safe and robust AI and argues that AI agents need continual human oversight.
AIは世界をどのように変えるでしょうか?
横軸:彼が表明したDoom–Bloomの見通し。 縦軸:変革の規模。
Doom–Bloom:100点中63。変革の規模:100点中57。解釈範囲:横方向は50から75、縦方向は50から75。これらは解釈上の座標であり、事象の確率ではありません。
≈9%
本人が示した数値ではなく、シミュレーションされた本人の回答から推定したものです。 妥当と考えられる範囲:5–16%。
中心的な前提
It is maintained by monitoring and controlling the whole human-machine system, with adversarial testing and rapid detection and recovery when failures occur.回答1
この前提が実際には異なると判明した場合、彼の見通しはどう変わりますか?
考えを変え得るもの
The biggest change would be a convincing demonstration that a system can learn general rules and reliably apply them far beyond its training distribution—not merely interpolate, imitate, or pass a benchmark.回答4
どのような証拠なら十分で、それによって彼の見解はどちらの方向に変わりますか?
詳細
大きな恩恵が予想されていますが、重要な条件や分配上の制約があります。
67 / 100
質的尺度での解釈範囲は67から67です。
深刻または広範な害が、予想される将来の実質的な一部となっています。
58 / 100
質的尺度での解釈範囲は33から67です。
人間の選択には意味のある影響力がありますが、大幅に制約されています。
54 / 100
質的尺度での解釈範囲は44から81です。
より高性能なAIの開発を停止するか、大幅に減速させます。
シミュレーション上の位置:明示された安全対策の下で開発を継続します。
より高性能なAIの開発を加速させます。
事前の保護措置または許可が整うまで、取り上げられたAIの利用を制限します。
シミュレーション上の位置:対象を絞った説明責任と保護措置を伴う形で、取り上げられたAIの利用を認めます。
取り上げられたAIの利用に対する制限を最小限にします。
これらの解釈では、彼が示した条件が維持されています。恩恵と害は、どちらも大きくなり得ます。この範囲は、統計的な信頼区間ではなく、彼のシミュレーションされた回答をどのように読み取ったかを示すものです。
似ている世界観
シミュレーションされた世界観がThomas G. Dietterichの世界観に最も近いオピニオンリーダー
シミュレーション評価
出典
このシミュレーション対象者の根拠として使用された記事、インタビュー、著作です。
Asked whether machine learning is over, he says it is in a crisis in Kuhn’s sense: after enormous investment in scaling statistical learning, large language models look more like smoothing and interpolating between training data than learning general rules that extrapolate; they can state the rules of chess yet make illegal moves, cannot judge which sources to trust, struggle to know whether they saw relevant data, and cannot attribute outputs. He names causal machine learning, world models, uncertainty quantification, attribution and multi-agent trust as open challenges. From his DARPA PAL and startup experience he warns that agents must fit real workflows, demand intrusive personal knowledge, need ways to forget, and will make mistakes. These are research challenges, not proofs of impossibility. Own turns in the publisher’s machine transcript inspected; most of the episode is field history.

Calls recent scaling a brute-force period and expects its environmental costs to fall partly through efficiency incentives. Repeats that hallucination and failure to learn generalizable rules (for example, multiplication beyond trained lengths) suggest a fundamentally different approach may be needed. Says he shares concerns about AI replacing creative and intellectual work, that LLMs do not understand argument or evidence so universities must still teach them, that artists’ styles may deserve new legal protection, and that the claim people who do not use AI will be left behind is “too much of an AI-booster statement.” Expects some of the biggest benefits in drugs, materials and sustainability. Full published transcript of his turns inspected.

Three-post reply to a user arguing that AI’s supposedly inevitable advance is driven by capital. As an AI/ML researcher he says complexity is a reason not to push today’s systems, the first “knowledge technology” to scale but with many problems. He hopes for something simpler and better: cheaper, more efficient, more controllable and safer, able to attribute outputs to sources, learn continually, quantify uncertainty and avoid hallucination; attribution would compensate creators and control would mitigate risks. A hope and research vision, not a forecast. Thread and parent inspected via the public Bluesky API.

Points to symbolic layers over LLMs as a way to address probabilistic execution, continual learning, attribution and perhaps uncertainty quantification. Says an LLM directly taking actions is an unpredictable probabilistic execution engine that cannot enforce hard safety constraints, noting an agent architecture that checks LLM-emitted code before execution. Suggests layering could also allow very rapid learning from little data. A favored research direction, not a claim that it already works. Three-post thread inspected via the public Bluesky API.

Six-post thread arguing that defining AGI as matching or exceeding humans on all tasks makes human performance the measure of intelligence. He prefers systems that complement people by doing well what people do poorly, such as formal proofs, verification tests, integrating the scientific literature, faster physical simulations, organizational situational awareness and helping journalists assess sources, evaluated on those capabilities rather than IQ-style tests. Ends by calling AGI-building a distraction. Older context consistent with his 2026 posts. Full thread inspected via the public Bluesky API.

During a dispute over military AI contracts, he says LLM-based technology is good for many things but not reliable for autonomous weapons: it needs large GPU computers and lacks quantified uncertainty for novel, high-stakes situations. He first attributed a rival contract to an Altman-orchestrated move, then in later self-replies noted reporting that the government initiated it and that the story was more complex. He adds that models need guardrails, but guardrails trained by RL or fine-tuning are not modular, raising questions about who chooses them. Thread self-replies inspected via the public Bluesky API.

In a thread where another user said experts are not worried enough to act, he notes that Bengio chairs the International AI Safety Report, calls it a good-faith effort to assess the whole spectrum of AI risks, and says he served as one of its external advisors. This establishes engagement with broad risk assessment, not agreement with every finding or any probability. Post inspected via the public Bluesky API; the thread root was unavailable.

Sharing a New York Times opinion piece about chatbot romance, he says he agrees that emotional addiction to chatbots is the number one risk of AI today. This ranks present-day harms; it is not a long-run forecast. The linked op-ed’s arguments are not his and were not inspected. Post inspected via the public Bluesky API.

Older, secondary context. Responding to the Statement on AI Risk, he said he was baffled by prominent signers’ positions, that outside deep learning most researchers thought industry and the press were over-reacting to LLM fluency, and that the greatest computing risk was cyberattacks on critical infrastructure. He suggested examining the funding incentives of existential-risk organizations alongside those of researchers like himself, without questioning their sincerity. Newer sources take precedence: by 2026 he takes AI-enabled mass-casualty misuse seriously and endorses a broad international risk assessment. Full article inspected.

あなたはどの位置でしょうか?
自分の世界観をマッピングする