Skip to main content

To trade in the US, go to polymarket.us

icon for Highest OpenAI score on Humanity’s Last Exam in 2026?

Highest OpenAI score on Humanity’s Last Exam in 2026?

icon for Highest OpenAI score on Humanity’s Last Exam in 2026?

Highest OpenAI score on Humanity’s Last Exam in 2026?

$107,820 Vol.

Dec 31, 2026
Polymarket

$107,820 Vol.

Polymarket

55%+

$34,352 Vol.

78%

60%+

$41,192 Vol.

33%

65%+

$9,431 Vol.

16%

70%+

$5,516 Vol.

10%

This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".OpenAI’s strongest reported HLE results sit at 54.7–58.7% as of mid-September 2026, trailing Anthropic’s Claude Fable 5.1 and Opus 5 configurations that lead public leaderboards at 59–65% with tool-augmented or high-effort reasoning setups. Rapid benchmark gains across frontier models this year—driven by improved chain-of-thought, verification loops, and agentic tool use—have compressed the gap, yet Anthropic maintains an edge in demonstrated expert-level reasoning on the 2,500-question academic benchmark. OpenAI’s next GPT variants or scaled reasoning systems before year-end represent the main swing factor, as historical release cadence and competitive pressure could push its top score into the low-to-mid 60s or leave it capped below 60%. Traders weigh these timelines against the December 31, 2026 resolution deadline.

This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No".

For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.

The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Volume
$107,820
End Date
Jan 1, 2027
Market Opened
Jul 23, 2026, 6:53 PM ET

Outcome proposed: Yes

No dispute

Final outcome: Yes

This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".OpenAI’s strongest reported HLE results sit at 54.7–58.7% as of mid-September 2026, trailing Anthropic’s Claude Fable 5.1 and Opus 5 configurations that lead public leaderboards at 59–65% with tool-augmented or high-effort reasoning setups. Rapid benchmark gains across frontier models this year—driven by improved chain-of-thought, verification loops, and agentic tool use—have compressed the gap, yet Anthropic maintains an edge in demonstrated expert-level reasoning on the 2,500-question academic benchmark. OpenAI’s next GPT variants or scaled reasoning systems before year-end represent the main swing factor, as historical release cadence and competitive pressure could push its top score into the low-to-mid 60s or leave it capped below 60%. Traders weigh these timelines against the December 31, 2026 resolution deadline.

This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No".

For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.

The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Volume
$107,820
End Date
Jan 1, 2027
Market Opened
Jul 23, 2026, 6:53 PM ET

Outcome proposed: Yes

No dispute

Final outcome: Yes

Beware of external links.

Frequently Asked Questions

"Highest OpenAI score on Humanity’s Last Exam in 2026?" is a prediction market on Polymarket with 5 possible outcomes where traders buy and sell shares based on what they believe will happen. The current leading outcome is "50%+" at 100%, followed by "55%+" at 78%. Prices reflect real-time crowd-sourced probabilities. For example, a share priced at 100¢ implies that the market collectively assigns a 100% chance to that outcome. These odds shift continuously as traders react to new developments and information. Shares in the correct outcome are redeemable for $1 each upon market resolution.

As of today, "Highest OpenAI score on Humanity’s Last Exam in 2026?" has generated $107.8K in total trading volume since the market launched on Jul 23, 2026. This level of trading activity reflects strong engagement from the Polymarket community and helps ensure that the current odds are informed by a deep pool of market participants. You can track live price movements and trade on any outcome directly on this page.

To trade on "Highest OpenAI score on Humanity’s Last Exam in 2026?," browse the 5 available outcomes listed on this page. Each outcome displays a current price representing the market's implied probability. To take a position, select the outcome you believe is most likely, choose "Yes" to trade in favor of it or "No" to trade against it, enter your amount, and click "Trade." If your chosen outcome is correct when the market resolves, your "Yes" shares pay out $1 each. If it's incorrect, they pay out $0. You can also sell your shares at any time before resolution if you want to lock in a profit or cut a loss.

The current frontrunner for "Highest OpenAI score on Humanity’s Last Exam in 2026?" is "50%+" at 100%, meaning the market assigns a 100% chance to that outcome. The next closest outcome is "55%+" at 78%. These odds update in real-time as traders buy and sell shares, so they reflect the latest collective view of what's most likely to happen. Check back frequently or bookmark this page to follow how the odds shift as new information emerges.

The resolution rules for "Highest OpenAI score on Humanity’s Last Exam in 2026?" define exactly what needs to happen for each outcome to be declared a winner — including the official data sources used to determine the result. You can review the complete resolution criteria in the "Rules" section on this page above the comments. We recommend reading the rules carefully before trading, as they specify the precise conditions, edge cases, and sources that govern how this market is settled.