Skip to main content

Để giao dịch tại Mỹ, hãy truy cập polymarket.us

icon for Highest OpenAI score on Humanity’s Last Exam in 2026?

Highest OpenAI score on Humanity’s Last Exam in 2026?

icon for Highest OpenAI score on Humanity’s Last Exam in 2026?

Highest OpenAI score on Humanity’s Last Exam in 2026?

$107,737 KL.

Dec 31, 2026
Polymarket

$107,737 KL.

Polymarket

55%+

$34,352 KL.

78%

60%+

$41,111 KL.

32%

65%+

$9,428 KL.

16%

70%+

$5,516 KL.

10%

This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".**Anthropic’s Claude Fable 5.1 and Opus 5 currently lead HLE leaderboards at 59–65% (text-only or tool-augmented variants as of mid-September 2026), while OpenAI’s strongest reported results sit at 54.7–58.7% (GPT-6 Astra, GPT-5.4 Pro, GPT-5.5 Pro).** This gap reflects Anthropic’s edge in adaptive reasoning, tool integration, and long-horizon task performance on the 2,500-question benchmark covering expert-level math, physics, biology, and other domains. Scores have risen rapidly in 2026 through better chain-of-thought, thinking modes, and agentic scaffolding, moving the frontier from low teens percent early in the year into the mid-50s. OpenAI’s releases have kept pace but trail slightly on the hardest closed-ended academic questions, where verification and calibration remain challenging. Trader consensus prices 50%+ and 55%+ thresholds very high, reflecting expected continued scaling and model iterations, while 65%+ and above carry lower odds due to the competitive lead and typical product timelines. Key upcoming catalysts include OpenAI’s next flagship releases or reasoning updates before year-end, potential tool-use or agent improvements, and any new independent evaluations. HLE remains unsaturated, so incremental gains from training advances or post-training techniques could still shift relative positioning quickly.

This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No".

For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.

The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Khối lượng
$107,737
Ngày kết thúc
Jan 1, 2027
Thị trường mở
Jul 23, 2026, 6:53 PM ET

Người giải quyết

0x65070BE91...
This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".**Anthropic’s Claude Fable 5.1 and Opus 5 currently lead HLE leaderboards at 59–65% (text-only or tool-augmented variants as of mid-September 2026), while OpenAI’s strongest reported results sit at 54.7–58.7% (GPT-6 Astra, GPT-5.4 Pro, GPT-5.5 Pro).** This gap reflects Anthropic’s edge in adaptive reasoning, tool integration, and long-horizon task performance on the 2,500-question benchmark covering expert-level math, physics, biology, and other domains. Scores have risen rapidly in 2026 through better chain-of-thought, thinking modes, and agentic scaffolding, moving the frontier from low teens percent early in the year into the mid-50s. OpenAI’s releases have kept pace but trail slightly on the hardest closed-ended academic questions, where verification and calibration remain challenging. Trader consensus prices 50%+ and 55%+ thresholds very high, reflecting expected continued scaling and model iterations, while 65%+ and above carry lower odds due to the competitive lead and typical product timelines. Key upcoming catalysts include OpenAI’s next flagship releases or reasoning updates before year-end, potential tool-use or agent improvements, and any new independent evaluations. HLE remains unsaturated, so incremental gains from training advances or post-training techniques could still shift relative positioning quickly.

This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No".

For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.

The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Khối lượng
$107,737
Ngày kết thúc
Jan 1, 2027
Thị trường mở
Jul 23, 2026, 6:53 PM ET

Người giải quyết

0x65070BE91...

Cẩn thận với liên kết bên ngoài.

Câu hỏi thường gặp

"Highest OpenAI score on Humanity’s Last Exam in 2026?" là thị trường dự đoán trên Polymarket với 5 kết quả có thể nơi các nhà giao dịch mua và bán cổ phần dựa trên điều họ tin sẽ xảy ra. Kết quả dẫn đầu hiện tại là "50%+" ở mức 100%, tiếp theo là "55%+" ở mức 78%. Giá phản ánh xác suất cộng đồng theo thời gian thực. Ví dụ, cổ phần ở giá 100¢ ngụ ý thị trường tập thể cho rằng có 100% khả năng cho kết quả đó. Tỷ lệ này thay đổi liên tục khi trader phản ứng với diễn biến và thông tin mới. Cổ phần đúng kết quả có thể đổi lấy $1 mỗi cổ phần khi thị trường được giải quyết.

Tính đến hôm nay, "Highest OpenAI score on Humanity’s Last Exam in 2026?" đã tạo $107.7K tổng khối lượng giao dịch kể từ khi thị trường mở vào Jul 23, 2026. Mức hoạt động giao dịch này phản ánh sự tham gia mạnh mẽ từ cộng đồng Polymarket và giúp đảm bảo tỷ lệ hiện tại được thông tin bởi nhóm người tham gia thị trường sâu rộng. Bạn có thể theo dõi biến động giá trực tiếp và giao dịch trên bất kỳ kết quả nào ngay trên trang này.

Để giao dịch trên "Highest OpenAI score on Humanity’s Last Exam in 2026?," duyệt 5 kết quả có sẵn trên trang này. Mỗi kết quả hiển thị giá hiện tại đại diện cho xác suất ngụ ý của thị trường. Để mở vị thế, chọn kết quả bạn tin là có khả năng nhất, chọn "Có" để giao dịch ủng hộ hoặc "Không" để giao dịch chống, nhập số tiền và nhấn "Giao dịch." Nếu kết quả bạn chọn đúng khi thị trường giải quyết, cổ phần "Có" của bạn trả $1 mỗi cổ phần. Nếu sai, chúng trả $0. Bạn cũng có thể bán cổ phần bất cứ lúc nào trước khi giải quyết nếu muốn chốt lời hoặc cắt lỗ.

Ứng viên dẫn đầu hiện tại cho "Highest OpenAI score on Humanity’s Last Exam in 2026?" là "50%+" ở mức 100%, nghĩa là thị trường cho 100% khả năng cho kết quả đó. Kết quả gần nhất tiếp theo là "55%+" ở mức 78%. Tỷ lệ cập nhật theo thời gian thực khi trader mua và bán cổ phần, phản ánh cái nhìn tập thể mới nhất về điều có khả năng xảy ra nhất. Kiểm tra thường xuyên hoặc đánh dấu trang này để theo dõi tỷ lệ thay đổi khi thông tin mới xuất hiện.

Quy tắc giải quyết cho "Highest OpenAI score on Humanity’s Last Exam in 2026?" định nghĩa chính xác điều gì cần xảy ra để mỗi kết quả được tuyên bố thắng — bao gồm nguồn dữ liệu chính thức được sử dụng để xác định kết quả. Bạn có thể xem tiêu chí giải quyết đầy đủ trong phần "Quy tắc" trên trang này phía trên bình luận. Chúng tôi khuyên đọc kỹ quy tắc trước khi giao dịch, vì chúng chỉ rõ điều kiện, trường hợp ngoại lệ và nguồn chính xác quản lý cách thị trường được thanh toán.