Anthropic’s Claude series currently leads Humanity’s Last Exam (HLE) leaderboards, with models such as Claude Opus 5 and Claude Fable 5 posting the highest verified scores in the 55–65% range on the 2,500-question graduate-level benchmark covering math, sciences, and reasoning. This positions traders to favor outcomes above 55% for the year’s peak Claude result, reflecting rapid capability gains through improved reasoning modes and scale. Competitive pressure from OpenAI’s GPT-5 variants (around 49–57%) and Google’s Gemini releases continues to drive iterative updates at Anthropic. Key swing factors for the remainder of 2026 include any new Claude releases, post-training enhancements, or tool-use optimizations before year-end that could push the annual high-water mark higher.
Экспериментальная сводка, созданная ИИ на основе данных Polymarket. Это не является торговой рекомендацией и не влияет на то, как разрешается этот рынок. · ОбновленоСамый высокий балл Клода на последнем экзамене человечества в 2026 году?
$60,471 Объем
55%+
86%
60%+
57%
65%+
16%
70%+
9%
75%+
5%
$60,471 Объем
55%+
86%
60%+
57%
65%+
16%
70%+
9%
75%+
5%
For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Открытие рынка: Jul 23, 2026, 6:42 PM ET
Источник определения исхода
https://agi.safe.ai/Resolver
0x65070BE91...For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Источник определения исхода
https://agi.safe.ai/Resolver
0x65070BE91...Anthropic’s Claude series currently leads Humanity’s Last Exam (HLE) leaderboards, with models such as Claude Opus 5 and Claude Fable 5 posting the highest verified scores in the 55–65% range on the 2,500-question graduate-level benchmark covering math, sciences, and reasoning. This positions traders to favor outcomes above 55% for the year’s peak Claude result, reflecting rapid capability gains through improved reasoning modes and scale. Competitive pressure from OpenAI’s GPT-5 variants (around 49–57%) and Google’s Gemini releases continues to drive iterative updates at Anthropic. Key swing factors for the remainder of 2026 include any new Claude releases, post-training enhancements, or tool-use optimizations before year-end that could push the annual high-water mark higher.
Экспериментальная сводка, созданная ИИ на основе данных Polymarket. Это не является торговой рекомендацией и не влияет на то, как разрешается этот рынок. · Обновлено



Не доверяй внешним ссылкам.
Не доверяй внешним ссылкам.
Часто задаваемые вопросы