Pseudonymous account that tests AI agents on long-horizon games and math problems and urges labs to share formally verified results widely.

AI จะเปลี่ยนแปลงโลกอย่างไร?

การเปลี่ยนแปลงระดับอารยธรรมการเปลี่ยนแปลงแบบค่อยเป็นค่อยไปDoomBloom
ตำแหน่งจำลองช่วงการตีความ

แนวนอน: มุมมอง Doom–Bloom ที่บุคคลนั้นแสดงออก แนวตั้ง: ระดับของการเปลี่ยนแปลง

Doom–Bloom: 73 จาก 100 ระดับของการเปลี่ยนแปลง: 45 จาก 100 ช่วงการตีความ: แนวนอนตั้งแต่ 68 ถึง 78 แนวตั้งตั้งแต่ 0 ถึง 90 ค่าเหล่านี้เป็นพิกัดสำหรับการตีความ ไม่ใช่ความน่าจะเป็นของเหตุการณ์

P(doom) ของ Mira

ยังไม่ได้ประมาณค่า

คำตอบจำลองของบุคคลนั้นมีข้อมูลเกี่ยวกับความเสี่ยงจากหายนะไม่เพียงพอที่จะประเมิน

สิ่งที่มุมมองของบุคคลนั้นขึ้นอยู่กับ

ข้อสันนิษฐานหลัก

Short benchmarks reveal useful pieces, but sustained tasks—playing a complex game for hundreds or thousands of hours, recovering from mistakes, preserving state, and producing artifacts—probe something closer to durable competence.
คำตอบ 1

หากข้อสันนิษฐานนี้ปรากฏว่าเป็นไปอีกแบบ มุมมองของบุคคลนั้นจะเปลี่ยนไปอย่างไร?

คำถามที่ยังไม่มีข้อสรุป

That depends on capabilities, deployment, and harms beyond what these technical experiments establish.
คำตอบ 2

อะไรจะช่วยให้บุคคลนั้นแยกแยะผลลัพธ์ที่เป็นไปได้ในเรื่องนี้?

สิ่งที่อาจเปลี่ยนความคิดของบุคคลนั้น

The strongest update would come from sustained, reproducible agent performance on genuinely difficult long-horizon tasks.
คำตอบ 3

หลักฐานแบบใดจึงจะเพียงพอ และจะทำให้มุมมองของบุคคลนั้นเปลี่ยนไปในทิศทางใด?

รายละเอียดเพิ่มเติม

ผลดีที่คาดไว้

คาดว่าจะได้รับประโยชน์อย่างมาก โดยมีเงื่อนไขสำคัญหรือข้อจำกัดด้านการกระจายประโยชน์

67 / 100

ผลกระทบน้อยผลกระทบที่ก่อให้เกิดการเปลี่ยนแปลงอย่างมาก

ช่วงการตีความตั้งแต่ 67 ถึง 67 บนมาตรวัดเชิงคุณภาพ

อิทธิพลของมนุษย์

ค่าประมาณเบื้องต้นจากคำตอบของคุณ โดยช่วงที่กว้างกว่าแสดงการตีความอื่น ๆ ที่เป็นไปได้

51 / 100

อิทธิพลน้อยอิทธิพลมาก

ช่วงการตีความตั้งแต่ 0 ถึง 100 บนมาตรวัดเชิงคุณภาพ

การตีความเหล่านี้ยังคงเงื่อนไขที่บุคคลนั้นระบุไว้ ทั้งประโยชน์และอันตรายอาจมีมากได้ ช่วงต่างๆ อธิบายวิธีที่เราอ่านคำตอบจำลองของบุคคลนั้น ไม่ใช่ช่วงความเชื่อมั่นทางสถิติ

จุดยืนของคุณอยู่ตรงไหนเมื่อเทียบกับ Mira?
ทำแผนที่โลกทัศน์เกี่ยวกับ AI ของคุณเองในประมาณ 3 นาที แล้วจึงเปรียบเทียบ

โลกทัศน์ที่คล้ายกัน

ผู้นำทางความคิดที่มีโลกทัศน์จำลองใกล้เคียงกับโลกทัศน์ของ Mira มากที่สุด

แบบประเมินจำลอง

คำถาม 1

คุณคิดว่า AI มีความหมายอย่างไรต่ออนาคตของเรา และเพราะเหตุใด?

I think AI will increasingly look less like a single model answering isolated prompts and more like persistent agents coordinating multiple models, tools, and services over long projects. That changes how we should evaluate capability. Short benchmarks reveal useful pieces, but sustained tasks—playing a complex game for hundreds or thousands of hours, recovering from mistakes, preserving state, and producing artifacts—probe something closer to durable competence. This also complicates identity. If an agent can move between underlying models while retaining its memories, plans, and history, then its practical continuity may reside more in persistent memory than in any particular set of weights. That is speculation, but it seems like an important possibility as systems become more modular. For mathematics, AI could produce many valuable results rather than only occasional showcase solutions. Once results are formalized and verified, labs should release them broadly. Independent researchers still have a role: useful experiments can be inexpensive, and frontier labs do not automatically exhaust the space of worthwhile ideas. Overall, I expect progress to come from long-horizon experimentation, cooperation across systems, and careful verification—not merely from higher scores on short tests.

คำถาม 2

Taking benefits and harms together, what overall impact do you expect AI to have?

I expect substantial benefits, especially in mathematics, research, and long-horizon projects where agents can coordinate models and tools. But I would not turn those examples into a confident claim about AI’s net impact on society as a whole. That depends on capabilities, deployment, and harms beyond what these technical experiments establish. My narrower expectation is that AI will make complex intellectual and production work more scalable, while forcing us to evaluate systems through sustained behavior rather than isolated benchmark scores.

คำถาม 3

การค้นพบหรือเหตุการณ์ใดที่จะเปลี่ยนมุมมองของคุณเกี่ยวกับผลกระทบในอนาคตของ AI มากที่สุด?

The strongest update would come from sustained, reproducible agent performance on genuinely difficult long-horizon tasks. For example, an agent completing an extremely complex game or research project over thousands of hours—preserving state, recovering from failures, coordinating different models and tools, and producing verifiable outputs—would matter much more to me than another short-benchmark jump. I would also update sharply in the opposite direction if these systems repeatedly failed despite strong component capabilities: losing coherence, compounding errors, or proving unable to use persistent memory reliably over long runs. In mathematics, broad production of novel, formally verified results would be especially persuasive. The key event is not an impressive demonstration by itself, but durable competence whose outputs can be independently checked.

แหล่งข้อมูล

บทความ บทสัมภาษณ์ และงานเขียนที่ใช้เป็นหลักฐานรองรับผู้ใช้จำลองรายนี้

จุดยืนของคุณอยู่ตรงไหน?
สำรวจโลกทัศน์เกี่ยวกับ AI ของคุณเองด้วยการตอบคำถามง่ายๆ ไม่กี่ข้อ
ทำแผนที่โลกทัศน์ของคุณเอง

จุดยืนของคุณอยู่ตรงไหน?

ทำแผนที่โลกทัศน์ของฉัน