Anthropic’s rapid rollout of its Claude 5 family, including the September 2026 launch of Fable 5.1, has positioned the company at the top of Humanity’s Last Exam leaderboards, with recent scores reaching 59–65% depending on evaluation protocols. The benchmark’s 2,500 expert-authored questions across advanced math, physics, biology, and other domains remain far from saturated, rewarding genuine reasoning over retrieval. OpenAI’s GPT-5/6 variants and models from Meta and Google trail by several points, underscoring Anthropic’s edge in frontier knowledge work. Traders are watching for any late-2026 model updates or tool-use refinements that could further lift Claude scores before year-end resolution.
สรุปจาก AI ทดลองที่อ้างอิงข้อมูลจาก Polymarket ไม่ใช่คำแนะนำในการเทรดและไม่มีผลต่อการตัดสินตลาดนี้ · อัปเดตแล้ว$109,253 ปริมาณ
55%+
92%
60%+
41%
65%+
19%
70%+
12%
75%+
9%
$109,253 ปริมาณ
55%+
92%
60%+
41%
65%+
19%
70%+
12%
75%+
9%
For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
ตลาดเปิดเมื่อ: Jul 23, 2026, 6:42 PM ET
แหล่งข้อมูลการตัดสินผล
https://agi.safe.ai/ผู้ตัดสินผล
0x65070BE91...For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
แหล่งข้อมูลการตัดสินผล
https://agi.safe.ai/ผู้ตัดสินผล
0x65070BE91...Anthropic’s rapid rollout of its Claude 5 family, including the September 2026 launch of Fable 5.1, has positioned the company at the top of Humanity’s Last Exam leaderboards, with recent scores reaching 59–65% depending on evaluation protocols. The benchmark’s 2,500 expert-authored questions across advanced math, physics, biology, and other domains remain far from saturated, rewarding genuine reasoning over retrieval. OpenAI’s GPT-5/6 variants and models from Meta and Google trail by several points, underscoring Anthropic’s edge in frontier knowledge work. Traders are watching for any late-2026 model updates or tool-use refinements that could further lift Claude scores before year-end resolution.
สรุปจาก AI ทดลองที่อ้างอิงข้อมูลจาก Polymarket ไม่ใช่คำแนะนำในการเทรดและไม่มีผลต่อการตัดสินตลาดนี้ · อัปเดตแล้ว



ระวังลิงก์ภายนอก
ระวังลิงก์ภายนอก
คำถามที่พบบ่อย