**Anthropic's latest Claude variants, including Fable 5.1, currently lead Humanity’s Last Exam leaderboards at 55–65% as of mid-September 2026, while the strongest reported Grok 4.x models sit at 42–44% on the same expert-authored 2,500-question benchmark of graduate-level knowledge and multi-step reasoning across math, sciences, and humanities.** This gap reflects xAI’s recent model releases focusing more on general capabilities and agentic features than on saturating this particular closed-book academic test, where tool use and protocol differences can shift reported scores. With only three months remaining in 2026, traders are pricing in the low probability of a Grok breakthrough that would close the multi-point deficit against faster-moving competitors like Anthropic and OpenAI. New Grok releases or evaluation methodology changes could still move the market, but historical iteration speeds make rapid gains to 50%+ unlikely without major verified progress.
Експериментальне резюме, згенероване ШІ з посиланням на дані Polymarket. Це не торгова порада і не впливає на вирішення цього ринку. · ОновленоHighest Grok score on Humanity’s Last Exam in 2026?
$118,567 Обс.
45%+
78%
50%+
51%
55%+
31%
60%+
19%
65%+
4%
$118,567 Обс.
45%+
78%
50%+
51%
55%+
31%
60%+
19%
65%+
4%
For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Ринок відкрито: Jul 23, 2026, 6:46 PM ET
Джерело вирішення
https://agi.safe.ai/Вирішувач
0x65070BE91...For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Джерело вирішення
https://agi.safe.ai/Вирішувач
0x65070BE91...**Anthropic's latest Claude variants, including Fable 5.1, currently lead Humanity’s Last Exam leaderboards at 55–65% as of mid-September 2026, while the strongest reported Grok 4.x models sit at 42–44% on the same expert-authored 2,500-question benchmark of graduate-level knowledge and multi-step reasoning across math, sciences, and humanities.** This gap reflects xAI’s recent model releases focusing more on general capabilities and agentic features than on saturating this particular closed-book academic test, where tool use and protocol differences can shift reported scores. With only three months remaining in 2026, traders are pricing in the low probability of a Grok breakthrough that would close the multi-point deficit against faster-moving competitors like Anthropic and OpenAI. New Grok releases or evaluation methodology changes could still move the market, but historical iteration speeds make rapid gains to 50%+ unlikely without major verified progress.
Експериментальне резюме, згенероване ШІ з посиланням на дані Polymarket. Це не торгова порада і не впливає на вирішення цього ринку. · Оновлено



Обережно з зовнішніми посиланнями.
Обережно з зовнішніми посиланнями.
Часті запитання