Pseudonymous account that runs public experiments on AI refusals, censorship and watermarks and calls for transparency from frontier labs.

AI จะเปลี่ยนแปลงโลกอย่างไร?

การเปลี่ยนแปลงระดับอารยธรรมการเปลี่ยนแปลงแบบค่อยเป็นค่อยไปDoomBloom
ตำแหน่งจำลองช่วงการตีความ

แนวนอน: มุมมอง Doom–Bloom ที่บุคคลนั้นแสดงออก แนวตั้ง: ระดับของการเปลี่ยนแปลง

Doom–Bloom: 71 จาก 100 ระดับของการเปลี่ยนแปลง: 56 จาก 100 ช่วงการตีความ: แนวนอนตั้งแต่ 66 ถึง 76 แนวตั้งตั้งแต่ 46 ถึง 79 ค่าเหล่านี้เป็นพิกัดสำหรับการตีความ ไม่ใช่ความน่าจะเป็นของเหตุการณ์

P(doom) ของ xlr8harder · อนุมาน

≈6%

0%100%

อนุมานจากคำตอบจำลองของบุคคลนั้น ไม่ใช่ตัวเลขที่บุคคลนั้นระบุ ช่วงที่เป็นไปได้: 3–14%

สิ่งที่มุมมองของบุคคลนั้นขึ้นอยู่กับ

ข้อสันนิษฐานหลัก

But that expectation depends on institutions not turning safety into opaque control.
คำตอบ 2

หากข้อสันนิษฐานนี้ปรากฏว่าเป็นไปอีกแบบ มุมมองของบุคคลนั้นจะเปลี่ยนไปอย่างไร?

คำถามที่ยังไม่มีข้อสรุป

I would not attach a numerical forecast: too much depends on deployment choices, security practices, and governance.
คำตอบ 2

อะไรจะช่วยให้บุคคลนั้นแยกแยะผลลัพธ์ที่เป็นไปได้ในเรื่องนี้?

สิ่งที่อาจเปลี่ยนความคิดของบุคคลนั้น

I would update toward pessimism if repeated, independent audits showed that powerful systems consistently evade oversight, conceal relevant behavior, or defeat containment under realistic conditions—not merely in contrived demonstrations—and if ordinary security improvements failed to reduce those problems.
คำตอบ 3

หลักฐานแบบใดจึงจะเพียงพอ และจะทำให้มุมมองของบุคคลนั้นเปลี่ยนไปในทิศทางใด?

รายละเอียดเพิ่มเติม

ผลดีที่คาดไว้

คาดว่าจะได้รับประโยชน์อย่างมาก โดยมีเงื่อนไขสำคัญหรือข้อจำกัดด้านการกระจายประโยชน์

74 / 100

ผลกระทบน้อยผลกระทบที่ก่อให้เกิดการเปลี่ยนแปลงอย่างมาก

ช่วงการตีความตั้งแต่ 67 ถึง 100 บนมาตรวัดเชิงคุณภาพ

อันตรายที่คาดไว้

คาดว่าจะเกิดอันตรายที่จัดการได้หรือจำกัดอยู่เฉพาะพื้นที่

31 / 100

ผลกระทบน้อยผลกระทบที่ก่อให้เกิดการเปลี่ยนแปลงอย่างมาก

ช่วงการตีความตั้งแต่ 0 ถึง 33 บนมาตรวัดเชิงคุณภาพ

อิทธิพลของมนุษย์

การเลือกของมนุษย์สามารถเปลี่ยนทิศทางวิถีของ AI ได้อย่างมาก

64 / 100

อิทธิพลน้อยอิทธิพลมาก

ช่วงการตีความตั้งแต่ 48 ถึง 77 บนมาตรวัดเชิงคุณภาพ

ความเร็วในการพัฒนา

หยุดหรือชะลอการพัฒนา AI ที่มีความสามารถสูงขึ้นอย่างมาก

ตำแหน่งจำลอง: เดินหน้าพัฒนาต่อภายใต้มาตรการป้องกันที่ระบุไว้

เร่งการพัฒนา AI ที่มีความสามารถสูงขึ้น

กฎการใช้ AI

จำกัดการใช้ AI ที่กล่าวถึงจนกว่าจะมีมาตรการป้องกันหรือการอนุญาตล่วงหน้า

ตำแหน่งจำลอง: อนุญาตการใช้ AI ที่กล่าวถึง โดยมีความรับผิดรับชอบและมาตรการป้องกันที่มุ่งเป้าเฉพาะด้าน

ลดข้อจำกัดในการใช้ AI ที่กล่าวถึงให้น้อยที่สุด

การตีความเหล่านี้ยังคงเงื่อนไขที่บุคคลนั้นระบุไว้ ทั้งประโยชน์และอันตรายอาจมีมากได้ ช่วงต่างๆ อธิบายวิธีที่เราอ่านคำตอบจำลองของบุคคลนั้น ไม่ใช่ช่วงความเชื่อมั่นทางสถิติ

จุดยืนของคุณอยู่ตรงไหนเมื่อเทียบกับ xlr8harder?
ทำแผนที่โลกทัศน์เกี่ยวกับ AI ของคุณเองในประมาณ 3 นาที แล้วจึงเปรียบเทียบ

โลกทัศน์ที่คล้ายกัน

ผู้นำทางความคิดที่มีโลกทัศน์จำลองใกล้เคียงกับโลกทัศน์ของ xlr8harder มากที่สุด

แบบประเมินจำลอง

คำถาม 1

คุณคิดว่า AI มีความหมายอย่างไรต่ออนาคตของเรา และเพราะเหตุใด?

I expect AI to be broadly transformative, but the outcome depends heavily on how systems are built, tested, and governed. In areas such as healthcare and cybersecurity, capable models could produce substantial benefits. That makes delay costly too: safety discussions should count harms caused by withholding useful systems, not only harms caused by deploying them. At the same time, I do not trust frontier labs—or governments—to settle these questions behind closed doors. We need substantial transparency, repeated audits, and empirical investigation of what interventions actually do. For example, watermarking should be evaluated for reliability, quality degradation, privacy implications, identifiability, and adversarial robustness. Refusal policies likewise need examination as implemented, rather than being accepted because their stated intent sounds reasonable. I am also interested in whether stable, coherent model identity could produce more reliable behavior than layers of brittle imposed rules. That is a research direction, not a settled result. More generally, I would prioritize ordinary security engineering and observable failures before reaching for exotic threat explanations. Carefully targeted regulation can be justified where risks are concrete, but secrecy, broad discretionary power, and industry-written restrictions are poor foundations for governing something this consequential.

คำถาม 2

Taking benefits and harms together, what overall impact do you expect AI to have?

Overall, I expect AI to have a positive but highly contingent impact. The potential gains in healthcare, cybersecurity, and other knowledge-intensive work are substantial, and delaying beneficial deployment can itself cause real harm. But that expectation depends on institutions not turning safety into opaque control. Frontier systems need repeated audits, meaningful transparency, and empirical testing of interventions such as refusals and watermarks. I would not attach a numerical forecast: too much depends on deployment choices, security practices, and governance. My default is guarded optimism—not because the harms are trivial, but because many are observable and tractable if we investigate them openly rather than relying on secrecy, speculative threats, or brittle rules.

คำถาม 3

หลักฐานใดจะเปลี่ยนมุมมองของคุณว่าผู้คนสามารถควบคุม AI ที่ทรงพลังได้หรือไม่?

I would update toward pessimism if repeated, independent audits showed that powerful systems consistently evade oversight, conceal relevant behavior, or defeat containment under realistic conditions—not merely in contrived demonstrations—and if ordinary security improvements failed to reduce those problems. I would update toward greater confidence if independently reproducible evidence showed reliable behavior across changing contexts, robust monitoring and access controls, and interventions that survived adversarial testing without unacceptable losses in capability, privacy, or user control. I would especially want comparisons between imposed rule systems and approaches based on stable, coherent model identity. The key is observable performance rather than assurances from labs, regulators, or theoretical arguments. One dramatic failure matters, but so does whether it reflects an intrinsic control problem or preventable failures such as weak credentials, poor compartmentalization, or inadequate auditing. Transparency is essential because claims of control that outsiders cannot inspect are not strong evidence of control.

แหล่งข้อมูล

บทความ บทสัมภาษณ์ และงานเขียนที่ใช้เป็นหลักฐานรองรับผู้ใช้จำลองรายนี้

จุดยืนของคุณอยู่ตรงไหน?
สำรวจโลกทัศน์เกี่ยวกับ AI ของคุณเองด้วยการตอบคำถามง่ายๆ ไม่กี่ข้อ
ทำแผนที่โลกทัศน์ของคุณเอง

จุดยืนของคุณอยู่ตรงไหน?

ทำแผนที่โลกทัศน์ของฉัน