Skip to main content

Untuk trading di AS, kunjungi polymarket.us

icon for Highest Grok score on Humanity’s Last Exam in 2026?

Highest Grok score on Humanity’s Last Exam in 2026?

icon for Highest Grok score on Humanity’s Last Exam in 2026?

Highest Grok score on Humanity’s Last Exam in 2026?

$119,543 Vol.

Dec 31, 2026
Polymarket

$119,543 Vol.

Polymarket

45%+

$44,627 Vol.

81%

50%+

$25,068 Vol.

41%

55%+

$29,262 Vol.

31%

60%+

$17,653 Vol.

19%

65%+

$2,934 Vol.

4%

This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".**Anthropic’s Claude Fable 5.1 and related Opus/Mythos 5 variants currently lead Humanity’s Last Exam leaderboards at 59–65% (Artificial Analysis and BenchLM snapshots as of mid-September 2026), while the strongest publicly reported Grok 4.6 scores sit near 42–43%.** Earlier 2025 Grok 4 releases posted roughly 25% closed-book and up to ~44–50% with tools or multi-agent setups, but xAI has not yet matched the top Anthropic or OpenAI configurations on the expert-authored 2,500-question benchmark. Trader focus centers on whether xAI’s ongoing larger training runs (including the recently referenced 2.5T+ Grok 4.8) and continued iteration through Q4 can push a Grok variant into the low-to-mid 50s or higher before year-end, especially on HLE-Rolling variants that incorporate community feedback. Key swing factors include the exact evaluation protocol (tools vs. closed-book, text-only vs. multimodal) and any late-year model releases or verified submissions.

This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No".

For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.

The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Volume
$119,543
Tanggal Berakhir
Jan 1, 2027
Pasar Dibuka
Jul 23, 2026, 6:46 PM ET

Sumber Resolusi

https://agi.safe.ai/
This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".**Anthropic’s Claude Fable 5.1 and related Opus/Mythos 5 variants currently lead Humanity’s Last Exam leaderboards at 59–65% (Artificial Analysis and BenchLM snapshots as of mid-September 2026), while the strongest publicly reported Grok 4.6 scores sit near 42–43%.** Earlier 2025 Grok 4 releases posted roughly 25% closed-book and up to ~44–50% with tools or multi-agent setups, but xAI has not yet matched the top Anthropic or OpenAI configurations on the expert-authored 2,500-question benchmark. Trader focus centers on whether xAI’s ongoing larger training runs (including the recently referenced 2.5T+ Grok 4.8) and continued iteration through Q4 can push a Grok variant into the low-to-mid 50s or higher before year-end, especially on HLE-Rolling variants that incorporate community feedback. Key swing factors include the exact evaluation protocol (tools vs. closed-book, text-only vs. multimodal) and any late-year model releases or verified submissions.

This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No".

For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.

The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
This market will resolve to "Yes" if any SpaceXAI Grok model achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Volume
$119,543
Tanggal Berakhir
Jan 1, 2027
Pasar Dibuka
Jul 23, 2026, 6:46 PM ET

Sumber Resolusi

https://agi.safe.ai/

Hati-hati dengan link eksternal.

Pertanyaan yang Sering Diajukan

"Highest Grok score on Humanity’s Last Exam in 2026?" adalah pasar prediksi di Polymarket dengan 5 hasil yang mungkin di mana trader membeli dan menjual saham berdasarkan apa yang mereka yakini akan terjadi. Hasil terdepan saat ini adalah "45%+" di 81%, diikuti oleh "50%+" di 41%. Harga mencerminkan probabilitas crowd-sourced real-time. Misalnya, saham yang dihargai 81¢ menyiratkan bahwa pasar secara kolektif memberikan peluang 81% pada hasil tersebut. Peluang ini bergeser terus-menerus saat trader bereaksi terhadap perkembangan dan informasi baru. Saham dengan hasil yang benar bisa ditukarkan seharga $1 setiap saham saat pasar diselesaikan.

Per hari ini, "Highest Grok score on Humanity’s Last Exam in 2026?" telah menghasilkan $119.5K dalam total volume trading sejak pasar diluncurkan pada Jul 23, 2026. Tingkat aktivitas trading ini mencerminkan keterlibatan kuat dari komunitas Polymarket dan membantu memastikan bahwa peluang saat ini diinformasikan oleh kumpulan besar peserta pasar. Kamu bisa melacak pergerakan harga langsung dan trading di hasil apa pun langsung di halaman ini.

Untuk trading di "Highest Grok score on Humanity’s Last Exam in 2026?," jelajahi 5 hasil yang tersedia di halaman ini. Setiap hasil menampilkan harga saat ini yang mewakili probabilitas tersirat pasar. Untuk mengambil posisi, pilih hasil yang menurutmu paling mungkin, pilih "Ya" untuk mendukungnya atau "Tidak" untuk menentangnya, masukkan jumlahmu, dan klik "Trade." Jika hasil pilihanmu benar saat pasar diselesaikan, saham "Ya" kamu membayar $1 masing-masing. Jika salah, mereka membayar $0. Kamu juga bisa menjual sahammu kapan saja sebelum resolusi jika kamu ingin mengamankan keuntungan atau memotong kerugian.

Unggulan saat ini untuk "Highest Grok score on Humanity’s Last Exam in 2026?" adalah "45%+" di 81%, yang berarti pasar memberikan peluang 81% pada hasil tersebut. Hasil terdekat berikutnya adalah "50%+" di 41%. Peluang ini diperbarui secara real-time saat trader membeli dan menjual saham, sehingga mencerminkan pandangan kolektif terbaru tentang apa yang paling mungkin terjadi. Cek kembali secara rutin atau tandai halaman ini untuk mengikuti bagaimana peluang bergeser saat informasi baru muncul.

Aturan resolusi untuk "Highest Grok score on Humanity’s Last Exam in 2026?" mendefinisikan dengan tepat apa yang harus terjadi agar setiap hasil dinyatakan sebagai pemenang — termasuk sumber data resmi yang digunakan untuk menentukan hasilnya. Kamu bisa meninjau kriteria resolusi lengkap di bagian "Aturan" di halaman ini di atas komentar. Kami menyarankan membaca aturan dengan cermat sebelum trading, karena mereka menentukan kondisi tepat, kasus khusus, dan sumber yang mengatur bagaimana pasar ini diselesaikan.