Recent official and independent evaluations of 2026 IMO problems show frontier large language models, including Claude Opus 5, GPT-5.6 variants, and Chinese systems like Huawei Celia and Xiaohongshu dots-note-3.0, achieving perfect 42/42 or gold-threshold scores of 30+/42 under timed, natural-language conditions. Open-weight pipelines such as NVIDIA Nemotron 3 Ultra and DeepSeek V4 Flash have also cleared the gold cutoff with reproducible inference methods. Despite these benchmarks, trader consensus heavily favors No at 86% implied probability because no AI has secured an official gold medal through formal IMO entry, judging, or the designated Grand Challenge resolution process. Key catalysts include ongoing debates over grading verification, lack of direct competition against human contestants, and strict resolution criteria that distinguish benchmark saturation from contest victory. Upcoming reports on official recognition could still shift sentiment before year-end.
基於Polymarket數據的AI實驗性摘要。這不是交易建議,也不影響該市場的結算方式。 · 更新於是
$23,316 交易量
$23,316 交易量
是
$23,316 交易量
$23,316 交易量
The resolution source is the IMO Grand Challenge (https://imo-grand-challenge.github.io/) and the Artificial Intelligence Math Olympiad (AIMO, https://aimoprize.com/). If either source demonstrates that an AI has won the challenge/prize before the resolution date, this market will resolve to "Yes".
市場開放時間: Nov 12, 2025, 5:08 PM ET
The resolution source is the IMO Grand Challenge (https://imo-grand-challenge.github.io/) and the Artificial Intelligence Math Olympiad (AIMO, https://aimoprize.com/). If either source demonstrates that an AI has won the challenge/prize before the resolution date, this market will resolve to "Yes".
Recent official and independent evaluations of 2026 IMO problems show frontier large language models, including Claude Opus 5, GPT-5.6 variants, and Chinese systems like Huawei Celia and Xiaohongshu dots-note-3.0, achieving perfect 42/42 or gold-threshold scores of 30+/42 under timed, natural-language conditions. Open-weight pipelines such as NVIDIA Nemotron 3 Ultra and DeepSeek V4 Flash have also cleared the gold cutoff with reproducible inference methods. Despite these benchmarks, trader consensus heavily favors No at 86% implied probability because no AI has secured an official gold medal through formal IMO entry, judging, or the designated Grand Challenge resolution process. Key catalysts include ongoing debates over grading verification, lack of direct competition against human contestants, and strict resolution criteria that distinguish benchmark saturation from contest victory. Upcoming reports on official recognition could still shift sentiment before year-end.
基於Polymarket數據的AI實驗性摘要。這不是交易建議,也不影響該市場的結算方式。 · 更新於



警惕外部連結哦。
警惕外部連結哦。
Frequently Asked Questions