คำถาม 1
Scott Alexander
x.com/slatestarcodexPsychiatrist and Astral Codex Ten blogger who sees large benefits and serious risks in AI and supports alignment research and negotiated slowdowns.
AI จะเปลี่ยนแปลงโลกอย่างไร?
แนวนอน: มุมมอง Doom–Bloom ที่เขาแสดงออก แนวตั้ง: ระดับของการเปลี่ยนแปลง
Doom–Bloom: 64 จาก 100 ระดับของการเปลี่ยนแปลง: 93 จาก 100 ช่วงการตีความ: แนวนอนตั้งแต่ 59 ถึง 75 แนวตั้งตั้งแต่ 88 ถึง 100 ค่าเหล่านี้เป็นพิกัดสำหรับการตีความ ไม่ใช่ความน่าจะเป็นของเหตุการณ์
20%
“I’m rounding both of them off to 20%.”
AI-caused human extinction, distinct from broader permanent curtailment of humanity’s future
My AI Opinions · มิ.ย. 2569
การทำงานและสถาบัน
My median forecast for AI able to perform roughly 90% of knowledge jobs is 2034.
คำตอบ 1
จัดกลุ่มตามหมุดหมายสำคัญ โดยไม่ได้เว้นระยะหรือเรียงตามวันที่ที่อนุมานไว้ AGI และ AI ที่เหนือกว่ามนุษย์ยังคงใช้คำนิยามของเขา
ข้อสันนิษฐานหลัก
The core concern is that systems trained through imperfect rewards may learn to deceive, exploit loopholes, or pursue objectives that diverge from ours once they become strategically capable.คำตอบ 1
หากข้อสันนิษฐานนี้ปรากฏว่าเป็นไปอีกแบบ มุมมองของเขาจะเปลี่ยนไปอย่างไร?
คำถามที่ยังไม่มีข้อสรุป
I’m uncertain about both.คำตอบ 1
อะไรจะช่วยให้เขาแยกแยะผลลัพธ์ที่เป็นไปได้ในเรื่องนี้?
สิ่งที่อาจเปลี่ยนความคิดของบุคคลนั้น
For example, repeated, adversarial demonstrations that highly capable systems remain honest and corrigible outside their training distribution—combined with interpretability that reveals why, rather than merely finding a reassuring-looking feature—would push my doom estimate substantially downward.คำตอบ 3
หลักฐานแบบใดจึงจะเพียงพอ และจะทำให้มุมมองของเขาเปลี่ยนไปในทิศทางใด?
รายละเอียดเพิ่มเติม
ยังมีการตีความที่เป็นไปได้หลายแบบ ได้แก่ คาดว่าจะได้รับประโยชน์ที่ก่อให้เกิดการเปลี่ยนแปลงและมีคุณค่าอย่างกว้างขวาง / คาดว่าจะได้รับประโยชน์อย่างมาก โดยมีเงื่อนไขสำคัญหรือข้อจำกัดด้านการกระจายประโยชน์
84 / 100
ช่วงการตีความตั้งแต่ 67 ถึง 100 บนมาตรวัดเชิงคุณภาพ
อันตรายร้ายแรงหรือแพร่หลายในวงกว้างเป็นส่วนสำคัญที่คาดว่าจะเกิดขึ้นในอนาคต
78 / 100
ช่วงการตีความตั้งแต่ 67 ถึง 100 บนมาตรวัดเชิงคุณภาพ
การเลือกของมนุษย์มีอิทธิพลอย่างมีนัยสำคัญ แต่ถูกจำกัดอย่างมาก
62 / 100
ช่วงการตีความตั้งแต่ 49 ถึง 76 บนมาตรวัดเชิงคุณภาพ
คาดว่า AI จะยังคงเป็นเครื่องมือที่มีขีดจำกัด
คาดว่า AI จะมีความสามารถทัดเทียมมนุษย์ในงานด้านการใช้ความคิดส่วนใหญ่
ตำแหน่งจำลอง: คาดว่า AI จะมีความสามารถเหนือกว่ามนุษย์อย่างมากในงานด้านการใช้ความคิด
การตีความเหล่านี้ยังคงเงื่อนไขที่เขาระบุไว้ ทั้งประโยชน์และอันตรายอาจมีมากได้ ช่วงต่างๆ อธิบายวิธีที่เราอ่านคำตอบจำลองของเขา ไม่ใช่ช่วงความเชื่อมั่นทางสถิติ
โลกทัศน์ที่คล้ายกัน
ผู้นำทางความคิดที่มีโลกทัศน์จำลองใกล้เคียงกับโลกทัศน์ของ Scott Alexander มากที่สุด
สิ่งที่ Scott Alexander เคยพูดเกี่ยวกับ AI
Scott Alexander writes that AI could bring large benefits and serious risks, and he supports alignment research and a negotiated slowdown.
“Plan A is still speculation, and still-speculative strong action is a perfectly reasonable response to still-speculative threats.”
Astral Codex Ten, AI Chip Regulation Is Not A Dystopian Surveillance State “The key insight is that if powerful AI is really as close and transformative as we think, then there’s a massive surplus that can satisfy everyone.”
Astral Codex Ten, Introducing Plan A “It’s increasingly clear that nobody has a plan for if this AI thing turns out to be real.”
Astral Codex Ten, Introducing Plan A “I find myself more optimistic about alignment than the average person who thinks about AI safety at all (although still more pessimistic than the average member of the population)”
Astral Codex Ten, My AI Opinions “A good pause strategy would involve both sides being able to monitor the other’s data centers to prevent illegal training”
Astral Codex Ten, My AI Opinions
อ้างอิงคำต่อคำจากแหล่งที่ลิงก์ไว้ ตรวจสอบเมื่อ 3 ต.ค. 2569
แบบประเมินจำลอง
แหล่งข้อมูล
บทความ บทสัมภาษณ์ และงานเขียนที่ใช้เป็นหลักฐานรองรับผู้ใช้จำลองรายนี้
His current first-person synthesis: AGI means ability to do 90% of knowledge jobs; median 2034, with uncertain research acceleration and diffusion. Reaffirms rounded 20% P(doom), with no fixed calendar deadline; broader permanent curtailment is separate. Supports both alignment research and negotiated slowing. Expects enormous postscarcity upside, but warns about dictatorship and human disempowerment.

Explains interpretability techniques and their limitations, including probes, sparse autoencoders, and activation verbalizers. Optimistic about useful practical investigation but rejects treating a detected feature or probe as a complete understanding or guaranteed safety solution.

Explicitly neutral about banning open weights now: values user ownership and freedom from corporate control, while expecting serious misuse difficulties. Prefers saving political capital for threats where warning shots may arrive too late. Distinguishes reactive policy opportunities for misuse from strategically concealed takeover.

Defends negotiated chip regulation and verifiable training limits against blanket claims of dystopia. Acknowledges real freedom costs, including future restrictions on new open-weight training, and risks that governments implement centralizing provisions without countervailing diffusion of power.

Introduces a proposed route to manage AI development while distributing power; criticizes vague calls merely to regulate more or less without specifying a desirable end state. Used as his attributed introduction and advocacy, not evidence that the scenario will occur.

Argues cheaper capable forecasting could improve institutional and personal decisions, yet worries people will ignore advice. Treats forecasting beyond human performance as a useful prospective test of the normal-technology view. Distinguishes anecdotes and startup claims from head-to-head competitions; admits resisting forecasts that challenge his own pause hopes.

Rejects the inference that requiring a new AI paradigm implies a safely distant AGI timeline. Argues paradigm changes can arrive soon and inherit existing compute infrastructure; wants explicit bottleneck arguments rather than reassurance by terminology.

Agrees growth cannot stay exponential forever but disputes placing the bend conveniently before dangerous capability. Demands a causal bottleneck model or a defensible forecasting prior instead of the slogan that all exponentials eventually flatten.

Satirical dialogue defends discussion of transparent, enforceable bilateral US-China slowing. Separates training limits from stopping existing inference, and legitimate negotiation or enforcement objections from falsely describing every pause proposal as unilateral.

Frames confident false answers as reward-shaped guessing rather than proof that AI cannot think. Treats the gap between trained reward and useful honest advice as an alignment issue; analogous human failures undermine easy dismissal of AI competence.

Separates training objectives from the representations and algorithms they produce, using evolution and human learning analogies. Argues next-token prediction does not itself establish that a system lacks reasoning or world models.

Identifies his part-time writing/publicity contribution and explicitly says the very fast scenario is not his median. Important provenance for his connection to AI Futures Project; use June 2026 personal forecasts instead of importing Daniel Kokotajlo’s timeline or scenario catastrophe probability.

จุดยืนของคุณอยู่ตรงไหน?
ทำแผนที่โลกทัศน์ของฉัน