Skip to main content

Aby handlować w USA, wejdź na polymarket.us

icon for Highest Grok score on Humanity’s Last Exam in 2026?

Highest Grok score on Humanity’s Last Exam in 2026?

icon for Highest Grok score on Humanity’s Last Exam in 2026?

Highest Grok score on Humanity’s Last Exam in 2026?

$115,656 Wol.

Dec 31, 2026
Polymarket

$115,656 Wol.

Polymarket

45%+

$44,039 Wol.

77%

50%+

$22,495 Wol.

47%

55%+

$28,676 Wol.

30%

60%+

$17,519 Wol.

22%

65%+

$2,928 Wol.

5%

This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".xAI’s Grok models currently trail on Humanity’s Last Exam (HLE), a 2,500-question expert benchmark of graduate-level knowledge and multi-step reasoning, with Grok 4.6 at 42.9% and 4.5 at 42.7% versus Claude Fable 5.1 leading at 55–65% on recent leaderboards. Trader sentiment centers on whether xAI’s aggressive scaling—evident in the ongoing 2.5-trillion-parameter Grok 4.8 training run and prior jumps like Grok 4 Heavy’s tool-augmented gains—can close the gap by year-end through larger models, refined reinforcement learning, and faster iteration cycles. Anthropic and OpenAI maintain leads via specialized reasoning variants, while HLE’s living updates and historical rapid progress (frontier rising from single digits to mid-40s% in under two years) create room for surprises. Key catalysts include the imminent 4.8 release, any new verified submissions before December 31, and potential tool-augmented or adaptive modes that could shift official leaderboard positions.

This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No".

For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.

The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Wolumen
$115,656
Data zakończenia
Jan 1, 2027
Rynek otwarty
Jul 23, 2026, 6:46 PM ET

Źródło rozstrzygnięcia

https://agi.safe.ai/

Rozstrzygający

0x65070BE91...
This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".xAI’s Grok models currently trail on Humanity’s Last Exam (HLE), a 2,500-question expert benchmark of graduate-level knowledge and multi-step reasoning, with Grok 4.6 at 42.9% and 4.5 at 42.7% versus Claude Fable 5.1 leading at 55–65% on recent leaderboards. Trader sentiment centers on whether xAI’s aggressive scaling—evident in the ongoing 2.5-trillion-parameter Grok 4.8 training run and prior jumps like Grok 4 Heavy’s tool-augmented gains—can close the gap by year-end through larger models, refined reinforcement learning, and faster iteration cycles. Anthropic and OpenAI maintain leads via specialized reasoning variants, while HLE’s living updates and historical rapid progress (frontier rising from single digits to mid-40s% in under two years) create room for surprises. Key catalysts include the imminent 4.8 release, any new verified submissions before December 31, and potential tool-augmented or adaptive modes that could shift official leaderboard positions.

This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No".

For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.

The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Wolumen
$115,656
Data zakończenia
Jan 1, 2027
Rynek otwarty
Jul 23, 2026, 6:46 PM ET

Źródło rozstrzygnięcia

https://agi.safe.ai/

Rozstrzygający

0x65070BE91...

Uważaj na linki zewnętrzne.

Często zadawane pytania

"Highest Grok score on Humanity’s Last Exam in 2026?" to rynek prognoz na Polymarket z 5 możliwymi wynikami, gdzie traderzy kupują i sprzedają udziały na podstawie tego, co ich zdaniem się wydarzy. Obecny wiodący wynik to "45%+" z 77%, za nim "50%+" z 47%. Ceny odzwierciedlają zbiorowe prawdopodobieństwa w czasie rzeczywistym. Na przykład udział wyceniony na 77¢ implikuje, że rynek zbiorowo przypisuje 77% szansy na ten wynik. Te kursy zmieniają się ciągle, gdy traderzy reagują na nowe informacje. Udziały w poprawnym wyniku można wymienić na $1 za sztukę po rozstrzygnięciu rynku.

Na dzień dzisiejszy "Highest Grok score on Humanity’s Last Exam in 2026?" wygenerował $115.7K łącznego wolumenu od uruchomienia rynku Jul 23, 2026. Ten poziom aktywności handlowej odzwierciedla silne zaangażowanie społeczności Polymarket i pomaga zapewnić, że bieżące kursy są informowane przez głęboką pulę uczestników rynku. Możesz śledzić ruchy cen na żywo i handlować na dowolny wynik bezpośrednio na tej stronie.

Aby handlować na "Highest Grok score on Humanity’s Last Exam in 2026?", przeglądaj 5 dostępnych wyników na tej stronie. Każdy wynik wyświetla bieżącą cenę reprezentującą implikowane prawdopodobieństwo rynku. Aby zająć pozycję, wybierz wynik, który uważasz za najbardziej prawdopodobny, wybierz "Tak", aby handlować na jego korzyść, lub "Nie", aby handlować przeciw niemu, wpisz kwotę i kliknij "Handluj". Jeśli wybrany wynik okaże się poprawny, Twoje udziały "Tak" wypłacą $1 za sztukę. Jeśli jest niepoprawny, wypłacą $0. Możesz też sprzedać swoje udziały w dowolnym momencie przed rozstrzygnięciem.

Obecnym faworytem dla "Highest Grok score on Humanity’s Last Exam in 2026?" jest "45%+" z 77%, co oznacza, że rynek przypisuje 77% szansy na ten wynik. Następny najbliższy wynik to "50%+" z 47%. Te kursy aktualizują się w czasie rzeczywistym, gdy traderzy kupują i sprzedają udziały, odzwierciedlając najnowszy zbiorowy pogląd na to, co jest najbardziej prawdopodobne. Sprawdzaj regularnie lub dodaj tę stronę do zakładek, aby śledzić zmiany kursów.

Zasady rozstrzygania "Highest Grok score on Humanity’s Last Exam in 2026?" określają dokładnie, co musi się wydarzyć, aby każdy wynik został ogłoszony zwycięzcą — w tym oficjalne źródła danych używane do ustalenia wyniku. Możesz przejrzeć pełne kryteria rozstrzygania w sekcji "Zasady" na tej stronie nad komentarzami. Zalecamy dokładne zapoznanie się z zasadami przed handlem, ponieważ określają one precyzyjne warunki, przypadki graniczne i źródła regulujące rozstrzyganie tego rynku.