Skip to main content
icon for AI model scores ≥ 90% on FrontierMath Benchmark before 2027?

AI model scores ≥ 90% on FrontierMath Benchmark before 2027?

icon for AI model scores ≥ 90% on FrontierMath Benchmark before 2027?

AI model scores ≥ 90% on FrontierMath Benchmark before 2027?

Dec 31

Dec 31

90% chance
Polymarket

$117,315 Vol.

90% chance
Polymarket

$117,315 Vol.

This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used. **Rapid gains in frontier AI reasoning systems have positioned models close to the 90% threshold on FrontierMath, driving the strong 90% market-implied odds for a breakthrough before 2027.** FrontierMath, developed by Epoch AI, consists of hundreds of original, unpublished research-level mathematics problems across tiers of increasing difficulty, designed to resist contamination and test expert-level capabilities that previously required hours or days of human mathematician effort. As of mid-2026, top models have advanced dramatically: Claude Fable 5 (max) reached 87.8% on the updated Tier 4 (v2) benchmark released in June 2026, while OpenAI’s GPT-5.6 Sol variants have posted scores in the low-to-mid 80s on the same set and up to 89% on legacy snapshots. These results reflect iterative improvements in long-horizon reasoning, chain-of-thought techniques, and domain-specific training for mathematical domains such as number theory and algebraic geometry. Traders see continued momentum from the ongoing release cycle of advanced large language models (LLMs) and reasoning systems through late 2026, including potential GPT-6 iterations or refined Claude and Gemini releases. Historical precedent shows scores rising from under 2% at the benchmark’s 2024 launch to over 40% on easier tiers and near 90% on harder ones within roughly 18–20 months. Key swing factors include further scaling of inference-time compute, algorithmic refinements, and possible specialized “AI co-mathematician” training runs. While timelines can slip and exact benchmark versions matter for resolution, the current trajectory and proximity to the threshold underpin the high consensus that 90% will be achieved before the end of 2026.

This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No".

The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.
Volume
$117,315
End Date
Dec 31, 2026
Market Opened
Nov 12, 2025, 5:15 PM ET
This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.
This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used. **Rapid gains in frontier AI reasoning systems have positioned models close to the 90% threshold on FrontierMath, driving the strong 90% market-implied odds for a breakthrough before 2027.** FrontierMath, developed by Epoch AI, consists of hundreds of original, unpublished research-level mathematics problems across tiers of increasing difficulty, designed to resist contamination and test expert-level capabilities that previously required hours or days of human mathematician effort. As of mid-2026, top models have advanced dramatically: Claude Fable 5 (max) reached 87.8% on the updated Tier 4 (v2) benchmark released in June 2026, while OpenAI’s GPT-5.6 Sol variants have posted scores in the low-to-mid 80s on the same set and up to 89% on legacy snapshots. These results reflect iterative improvements in long-horizon reasoning, chain-of-thought techniques, and domain-specific training for mathematical domains such as number theory and algebraic geometry. Traders see continued momentum from the ongoing release cycle of advanced large language models (LLMs) and reasoning systems through late 2026, including potential GPT-6 iterations or refined Claude and Gemini releases. Historical precedent shows scores rising from under 2% at the benchmark’s 2024 launch to over 40% on easier tiers and near 90% on harder ones within roughly 18–20 months. Key swing factors include further scaling of inference-time compute, algorithmic refinements, and possible specialized “AI co-mathematician” training runs. While timelines can slip and exact benchmark versions matter for resolution, the current trajectory and proximity to the threshold underpin the high consensus that 90% will be achieved before the end of 2026.

This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No".

The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.
Volume
$117,315
End Date
Dec 31, 2026
Market Opened
Nov 12, 2025, 5:15 PM ET
This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.

Beware of external links.

Frequently Asked Questions

"AI model scores ≥ 90% on FrontierMath Benchmark before 2027?" is a prediction market on Polymarket where traders buy and sell "Yes" or "No" shares based on whether they believe this event will happen. The current crowd-sourced probability is 90% for "Yes." For example, if "Yes" is priced at 90¢, the market collectively assigns a 90% chance that this event will occur. These odds shift continuously as traders react to new developments and information. Shares in the correct outcome are redeemable for $1 each upon market resolution.

As of today, "AI model scores ≥ 90% on FrontierMath Benchmark before 2027?" has generated $117.3K in total trading volume since the market launched on Nov 12, 2025. This level of trading activity reflects strong engagement from the Polymarket community and helps ensure that the current odds are informed by a deep pool of market participants. You can track live price movements and trade on any outcome directly on this page.

To trade on "AI model scores ≥ 90% on FrontierMath Benchmark before 2027?," simply choose whether you believe the answer is "Yes" or "No." Each side has a current price that reflects the market's implied probability. Enter your amount and click "Trade." If you buy "Yes" shares and the outcome resolves as "Yes," each share pays out $1. If it resolves as "No," your "Yes" shares pay $0. You can also sell your shares at any time before resolution if you want to lock in a profit or cut a loss.

The current probability for "AI model scores ≥ 90% on FrontierMath Benchmark before 2027?" is 90% for "Yes." This means the Polymarket crowd currently believes there is a 90% chance that this event will occur. These odds update in real-time based on actual trades, providing a continuously updated signal of what the market expects to happen.

The resolution rules for "AI model scores ≥ 90% on FrontierMath Benchmark before 2027?" define exactly what needs to happen for each outcome to be declared a winner — including the official data sources used to determine the result. You can review the complete resolution criteria in the "Rules" section on this page above the comments. We recommend reading the rules carefully before trading, as they specify the precise conditions, edge cases, and sources that govern how this market is settled.