Skip to main content

Para negociar nos EUA, acesse polymarket.us

icon for Maior pontuação de Grok no Último Exame da Humanidade em 2026?

Maior pontuação de Grok no Último Exame da Humanidade em 2026?

icon for Maior pontuação de Grok no Último Exame da Humanidade em 2026?

Maior pontuação de Grok no Último Exame da Humanidade em 2026?

$119,356 Vol.

31 dez 2026
Polymarket

$119,356 Vol.

Polymarket

45%+

$44,627 Vol.

81%

50%+

$24,881 Vol.

41%

55%+

$29,262 Vol.

31%

60%+

$17,653 Vol.

19%

65%+

$2,934 Vol.

4%

This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".**Anthropic’s Claude Fable 5.1 and related Opus/Mythos 5 variants currently lead Humanity’s Last Exam leaderboards at 59–65% (Artificial Analysis and BenchLM snapshots as of mid-September 2026), while the strongest publicly reported Grok 4.6 scores sit near 42–43%.** Earlier 2025 Grok 4 releases posted roughly 25% closed-book and up to ~44–50% with tools or multi-agent setups, but xAI has not yet matched the top Anthropic or OpenAI configurations on the expert-authored 2,500-question benchmark. Trader focus centers on whether xAI’s ongoing larger training runs (including the recently referenced 2.5T+ Grok 4.8) and continued iteration through Q4 can push a Grok variant into the low-to-mid 50s or higher before year-end, especially on HLE-Rolling variants that incorporate community feedback. Key swing factors include the exact evaluation protocol (tools vs. closed-book, text-only vs. multimodal) and any late-year model releases or verified submissions.

This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No".

For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.

The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Volume
$119,356
Data de Término
1 jan 2027
Mercado Aberto
Jul 23, 2026, 6:46 PM ET

Fonte de resolução

https://agi.safe.ai/
This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".**Anthropic’s Claude Fable 5.1 and related Opus/Mythos 5 variants currently lead Humanity’s Last Exam leaderboards at 59–65% (Artificial Analysis and BenchLM snapshots as of mid-September 2026), while the strongest publicly reported Grok 4.6 scores sit near 42–43%.** Earlier 2025 Grok 4 releases posted roughly 25% closed-book and up to ~44–50% with tools or multi-agent setups, but xAI has not yet matched the top Anthropic or OpenAI configurations on the expert-authored 2,500-question benchmark. Trader focus centers on whether xAI’s ongoing larger training runs (including the recently referenced 2.5T+ Grok 4.8) and continued iteration through Q4 can push a Grok variant into the low-to-mid 50s or higher before year-end, especially on HLE-Rolling variants that incorporate community feedback. Key swing factors include the exact evaluation protocol (tools vs. closed-book, text-only vs. multimodal) and any late-year model releases or verified submissions.

This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No".

For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.

The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Volume
$119,356
Data de Término
1 jan 2027
Mercado Aberto
Jul 23, 2026, 6:46 PM ET

Fonte de resolução

https://agi.safe.ai/

Cuidado com os links externos.

Frequently Asked Questions

"Maior pontuação de Grok no Último Exame da Humanidade em 2026?" is a prediction market on Polymarket with 5 possible outcomes where traders buy and sell shares based on what they believe will happen. The current leading outcome is "45%+" at 81%, followed by "50%+" at 41%. Prices reflect real-time crowd-sourced probabilities. For example, a share priced at 81¢ implies that the market collectively assigns a 81% chance to that outcome. These odds shift continuously as traders react to new developments and information. Shares in the correct outcome are redeemable for $1 each upon market resolution.

As of today, "Maior pontuação de Grok no Último Exame da Humanidade em 2026?" has generated $119.4K in total trading volume since the market launched on Jul 23, 2026. This level of trading activity reflects strong engagement from the Polymarket community and helps ensure that the current odds are informed by a deep pool of market participants. You can track live price movements and trade on any outcome directly on this page.

To trade on "Maior pontuação de Grok no Último Exame da Humanidade em 2026?," browse the 5 available outcomes listed on this page. Each outcome displays a current price representing the market's implied probability. To take a position, select the outcome you believe is most likely, choose "Yes" to trade in favor of it or "No" to trade against it, enter your amount, and click "Trade." If your chosen outcome is correct when the market resolves, your "Yes" shares pay out $1 each. If it's incorrect, they pay out $0. You can also sell your shares at any time before resolution if you want to lock in a profit or cut a loss.

The current frontrunner for "Maior pontuação de Grok no Último Exame da Humanidade em 2026?" is "45%+" at 81%, meaning the market assigns a 81% chance to that outcome. The next closest outcome is "50%+" at 41%. These odds update in real-time as traders buy and sell shares, so they reflect the latest collective view of what's most likely to happen. Check back frequently or bookmark this page to follow how the odds shift as new information emerges.

The resolution rules for "Maior pontuação de Grok no Último Exame da Humanidade em 2026?" define exactly what needs to happen for each outcome to be declared a winner — including the official data sources used to determine the result. You can review the complete resolution criteria in the "Rules" section on this page above the comments. We recommend reading the rules carefully before trading, as they specify the precise conditions, edge cases, and sources that govern how this market is settled.