Meta’s Muse Spark 1.1 model has posted scores near 62% on Humanity’s Last Exam leaderboards as of September 2026, placing it within a few points of Anthropic’s leading Claude Fable 5.1 and Opus 5 variants at 64.7–65%. This performance reflects Meta’s rapid iteration on adaptive reasoning techniques, expanded training infrastructure, and targeted post-training, enabling competitive results on the 2,500-question expert-crafted benchmark that spans advanced mathematics, physics, and humanities while resisting memorization. Traders are monitoring whether Meta can close the remaining gap or surpass 65% before year-end through anticipated model drops or scaling gains, though Anthropic’s consistent edge and the benchmark’s deliberate difficulty introduce uncertainty around exact thresholds.
สรุปจาก AI ทดลองที่อ้างอิงข้อมูลจาก Polymarket ไม่ใช่คำแนะนำในการเทรดและไม่มีผลต่อการตัดสินตลาดนี้ · อัปเดตแล้ว$58,346 ปริมาณ
55%+
43%
60%+
28%
65%+
22%
70%+
8%
$58,346 ปริมาณ
55%+
43%
60%+
28%
65%+
22%
70%+
8%
For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
ตลาดเปิดเมื่อ: Jul 23, 2026, 6:48 PM ET
แหล่งข้อมูลการตัดสินผล
https://agi.safe.ai/ผู้ตัดสินผล
0x65070BE91...For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
แหล่งข้อมูลการตัดสินผล
https://agi.safe.ai/ผู้ตัดสินผล
0x65070BE91...Meta’s Muse Spark 1.1 model has posted scores near 62% on Humanity’s Last Exam leaderboards as of September 2026, placing it within a few points of Anthropic’s leading Claude Fable 5.1 and Opus 5 variants at 64.7–65%. This performance reflects Meta’s rapid iteration on adaptive reasoning techniques, expanded training infrastructure, and targeted post-training, enabling competitive results on the 2,500-question expert-crafted benchmark that spans advanced mathematics, physics, and humanities while resisting memorization. Traders are monitoring whether Meta can close the remaining gap or surpass 65% before year-end through anticipated model drops or scaling gains, though Anthropic’s consistent edge and the benchmark’s deliberate difficulty introduce uncertainty around exact thresholds.
สรุปจาก AI ทดลองที่อ้างอิงข้อมูลจาก Polymarket ไม่ใช่คำแนะนำในการเทรดและไม่มีผลต่อการตัดสินตลาดนี้ · อัปเดตแล้ว



ระวังลิงก์ภายนอก
ระวังลิงก์ภายนอก
คำถามที่พบบ่อย