**Anthropic's latest Claude variants, including Fable 5.1, currently lead Humanity’s Last Exam leaderboards at 55–65% as of mid-September 2026, while the strongest reported Grok 4.x models sit at 42–44% on the same expert-authored 2,500-question benchmark of graduate-level knowledge and multi-step reasoning across math, sciences, and humanities.** This gap reflects xAI’s recent model releases focusing more on general capabilities and agentic features than on saturating this particular closed-book academic test, where tool use and protocol differences can shift reported scores. With only three months remaining in 2026, traders are pricing in the low probability of a Grok breakthrough that would close the multi-point deficit against faster-moving competitors like Anthropic and OpenAI. New Grok releases or evaluation methodology changes could still move the market, but historical iteration speeds make rapid gains to 50%+ unlikely without major verified progress.
Экспериментальная сводка, созданная ИИ на основе данных Polymarket. Это не является торговой рекомендацией и не влияет на то, как разрешается этот рынок. · ОбновленоСамый высокий балл Grok на последнем экзамене человечества в 2026 году?
$118,419 Объем
45%+
78%
50%+
51%
55%+
31%
60%+
19%
65%+
4%
$118,419 Объем
45%+
78%
50%+
51%
55%+
31%
60%+
19%
65%+
4%
For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Открытие рынка: Jul 23, 2026, 6:46 PM ET
Источник определения исхода
https://agi.safe.ai/Кто определяет исход
0x65070BE91...For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Источник определения исхода
https://agi.safe.ai/Кто определяет исход
0x65070BE91...**Anthropic's latest Claude variants, including Fable 5.1, currently lead Humanity’s Last Exam leaderboards at 55–65% as of mid-September 2026, while the strongest reported Grok 4.x models sit at 42–44% on the same expert-authored 2,500-question benchmark of graduate-level knowledge and multi-step reasoning across math, sciences, and humanities.** This gap reflects xAI’s recent model releases focusing more on general capabilities and agentic features than on saturating this particular closed-book academic test, where tool use and protocol differences can shift reported scores. With only three months remaining in 2026, traders are pricing in the low probability of a Grok breakthrough that would close the multi-point deficit against faster-moving competitors like Anthropic and OpenAI. New Grok releases or evaluation methodology changes could still move the market, but historical iteration speeds make rapid gains to 50%+ unlikely without major verified progress.
Экспериментальная сводка, созданная ИИ на основе данных Polymarket. Это не является торговой рекомендацией и не влияет на то, как разрешается этот рынок. · Обновлено



Не доверяй внешним ссылкам.
Не доверяй внешним ссылкам.
Часто задаваемые вопросы