Skip to main content
icon for 2026年人类最后一次考试的最高OpenAI分数?

2026年人类最后一次考试的最高OpenAI分数?

icon for 2026年人类最后一次考试的最高OpenAI分数?

2026年人类最后一次考试的最高OpenAI分数?

最新
2026-12-31
Polymarket

$1,875 交易量

Polymarket

50%及以上

$68 交易量

59%

55%及以上

$805 交易量

42%

60%及以上

$670 交易量

14%

65%以上

$0 交易量

26%

70%及以上

$332 交易量

8%

This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".OpenAI trails current Humanity’s Last Exam leaderboards, with its strongest GPT-5.6 variants scoring around 47-57% while Anthropic’s Claude models (Fable 5, Mythos Preview) hold top spots near 53-65%. The 2,500-question expert benchmark, released in 2025 by the Center for AI Safety and Scale AI, tests graduate-level knowledge across math, sciences, and humanities and has proven resistant to rapid gains beyond agentic techniques like OpenAI’s earlier Deep Research system. Traders focus on OpenAI’s history of iterative releases, substantial compute resources, and potential 2026 model launches or fine-tuning advances that could close the gap, balanced against sustained competition from Anthropic and Google DeepMind. Key catalysts include any major GPT updates or benchmark-specific optimizations before year-end resolution.

This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No".

For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.

The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
交易量
$1,875
结束日期
2026-12-31
市场开放时间
Jul 23, 2026, 6:53 PM ET
This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".OpenAI trails current Humanity’s Last Exam leaderboards, with its strongest GPT-5.6 variants scoring around 47-57% while Anthropic’s Claude models (Fable 5, Mythos Preview) hold top spots near 53-65%. The 2,500-question expert benchmark, released in 2025 by the Center for AI Safety and Scale AI, tests graduate-level knowledge across math, sciences, and humanities and has proven resistant to rapid gains beyond agentic techniques like OpenAI’s earlier Deep Research system. Traders focus on OpenAI’s history of iterative releases, substantial compute resources, and potential 2026 model launches or fine-tuning advances that could close the gap, balanced against sustained competition from Anthropic and Google DeepMind. Key catalysts include any major GPT updates or benchmark-specific optimizations before year-end resolution.

This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No".

For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.

The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
交易量
$1,875
结束日期
2026-12-31
市场开放时间
Jul 23, 2026, 6:53 PM ET
This market will resolve to "Yes" if any model published by OpenAI achieves at least the specified accuracy on Humanity’s Last Exam by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric. The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".

警惕外部链接哦。

常见问题

"2026年人类最后一次考试的最高OpenAI分数?"是 Polymarket 上一个拥有 5 个可能结果的预测市场,交易者根据自己的判断买卖份额。当前领先结果为"50%及以上",概率为 59%,其次是"55%及以上",概率为 42%。价格反映社区的实时概率。例如,价格为 59¢ 的份额意味着市场集体认为该结果的概率为 59%。这些赔率会随着交易者的反应而不断变化。正确结果的份额在市场结算时可兑换为每份 $1。

"2026年人类最后一次考试的最高OpenAI分数?"是 Polymarket 上新创建的市场,于Jul 23, 2026上线。作为一个新市场,这是你率先设定赔率并建立初始价格信号的机会。你也可以将本页加入书签,以便跟踪交易量和活动。

要在"2026年人类最后一次考试的最高OpenAI分数?"上交易,浏览本页上列出的 5 个可用结果。每个结果显示一个代表市场隐含概率的当前价格。要建仓,选择你认为最可能的结果,选择"是"支持或"否"反对,输入金额并点击"交易"。如果你选择的结果在市场结算时正确,你的"是"份额每份支付 $1。如果不正确,支付 $0。你也可以在结算前随时卖出份额。

"2026年人类最后一次考试的最高OpenAI分数?"的当前领先者是"50%及以上",概率为 59%,意味着市场对该结果的概率评估为 59%。紧随其后的结果是"55%及以上",概率为 42%。这些赔率随着交易者买卖份额而实时更新。请经常回来查看或将本页加入书签。

"2026年人类最后一次考试的最高OpenAI分数?"的结算规则明确定义了每个结果被宣布为获胜者所需满足的条件——包括用于确定结果的官方数据来源。你可以在本页评论上方的"规则"部分查看完整的结算标准。我们建议在交易前仔细阅读规则,因为它们规定了精确的条件、特殊情况和数据来源。