Skip to main content
icon for ¿El modelo de IA obtiene un puntaje ≥ 90% en FrontierMath Benchmark antes de 2027?

¿El modelo de IA obtiene un puntaje ≥ 90% en FrontierMath Benchmark antes de 2027?

icon for ¿El modelo de IA obtiene un puntaje ≥ 90% en FrontierMath Benchmark antes de 2027?

¿El modelo de IA obtiene un puntaje ≥ 90% en FrontierMath Benchmark antes de 2027?

dic 31

dic 31

90% probabilidad
Polymarket

$117,315 Vol.

90% probabilidad
Polymarket

$117,315 Vol.

This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used. **Rapid gains in frontier AI reasoning systems have positioned models close to the 90% threshold on FrontierMath, driving the strong 90% market-implied odds for a breakthrough before 2027.** FrontierMath, developed by Epoch AI, consists of hundreds of original, unpublished research-level mathematics problems across tiers of increasing difficulty, designed to resist contamination and test expert-level capabilities that previously required hours or days of human mathematician effort. As of mid-2026, top models have advanced dramatically: Claude Fable 5 (max) reached 87.8% on the updated Tier 4 (v2) benchmark released in June 2026, while OpenAI’s GPT-5.6 Sol variants have posted scores in the low-to-mid 80s on the same set and up to 89% on legacy snapshots. These results reflect iterative improvements in long-horizon reasoning, chain-of-thought techniques, and domain-specific training for mathematical domains such as number theory and algebraic geometry. Traders see continued momentum from the ongoing release cycle of advanced large language models (LLMs) and reasoning systems through late 2026, including potential GPT-6 iterations or refined Claude and Gemini releases. Historical precedent shows scores rising from under 2% at the benchmark’s 2024 launch to over 40% on easier tiers and near 90% on harder ones within roughly 18–20 months. Key swing factors include further scaling of inference-time compute, algorithmic refinements, and possible specialized “AI co-mathematician” training runs. While timelines can slip and exact benchmark versions matter for resolution, the current trajectory and proximity to the threshold underpin the high consensus that 90% will be achieved before the end of 2026.

This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No".

The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.
Volumen
$117,315
Fecha de finalización
31 dic 2026
Mercado abierto
Nov 12, 2025, 5:15 PM ET
This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.
This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used. **Rapid gains in frontier AI reasoning systems have positioned models close to the 90% threshold on FrontierMath, driving the strong 90% market-implied odds for a breakthrough before 2027.** FrontierMath, developed by Epoch AI, consists of hundreds of original, unpublished research-level mathematics problems across tiers of increasing difficulty, designed to resist contamination and test expert-level capabilities that previously required hours or days of human mathematician effort. As of mid-2026, top models have advanced dramatically: Claude Fable 5 (max) reached 87.8% on the updated Tier 4 (v2) benchmark released in June 2026, while OpenAI’s GPT-5.6 Sol variants have posted scores in the low-to-mid 80s on the same set and up to 89% on legacy snapshots. These results reflect iterative improvements in long-horizon reasoning, chain-of-thought techniques, and domain-specific training for mathematical domains such as number theory and algebraic geometry. Traders see continued momentum from the ongoing release cycle of advanced large language models (LLMs) and reasoning systems through late 2026, including potential GPT-6 iterations or refined Claude and Gemini releases. Historical precedent shows scores rising from under 2% at the benchmark’s 2024 launch to over 40% on easier tiers and near 90% on harder ones within roughly 18–20 months. Key swing factors include further scaling of inference-time compute, algorithmic refinements, and possible specialized “AI co-mathematician” training runs. While timelines can slip and exact benchmark versions matter for resolution, the current trajectory and proximity to the threshold underpin the high consensus that 90% will be achieved before the end of 2026.

This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No".

The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.
Volumen
$117,315
Fecha de finalización
31 dic 2026
Mercado abierto
Nov 12, 2025, 5:15 PM ET
This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.

Cuidado con los enlaces externos.

Preguntas frecuentes

"¿El modelo de IA obtiene un puntaje ≥ 90% en FrontierMath Benchmark antes de 2027?" es un mercado de predicción en Polymarket con 2 resultados posibles donde los operadores compran y venden acciones según lo que creen que sucederá. El resultado líder actual es "¿Modelos de IA puntúan ≥ 90% en el FrontMath Benchmark antes de 2027?" con 90%. Los precios reflejan probabilidades en tiempo real de la comunidad. Por ejemplo, una acción cotizada a 90¢ implica que el mercado colectivamente asigna una probabilidad de 90% a ese resultado. Estas probabilidades cambian continuamente a medida que los operadores reaccionan a nuevos desarrollos. Las acciones del resultado correcto son canjeables por $1 cada una tras la resolución del mercado.

A día de hoy, "¿El modelo de IA obtiene un puntaje ≥ 90% en FrontierMath Benchmark antes de 2027?" ha generado $117.3K en volumen total de trading desde que el mercado se lanzó el Nov 12, 2025. Este nivel de actividad refleja un fuerte compromiso de la comunidad de Polymarket y ayuda a garantizar que las probabilidades actuales estén respaldadas por un amplio grupo de participantes del mercado. Puedes seguir los movimientos de precios en vivo y operar en cualquier resultado directamente en esta página.

Para operar en "¿El modelo de IA obtiene un puntaje ≥ 90% en FrontierMath Benchmark antes de 2027?", explora los 2 resultados disponibles en esta página. Cada resultado muestra un precio actual que representa la probabilidad implícita del mercado. Para tomar una posición, selecciona el resultado que consideres más probable, elige "Sí" para operar a favor o "No" para operar en contra, introduce tu cantidad y haz clic en "Operar". Si tu resultado elegido es correcto cuando el mercado se resuelve, tus acciones de "Sí" pagan $1 cada una. Si es incorrecto, pagan $0. También puedes vender tus acciones en cualquier momento antes de la resolución.

El favorito actual para "¿El modelo de IA obtiene un puntaje ≥ 90% en FrontierMath Benchmark antes de 2027?" es "¿Modelos de IA puntúan ≥ 90% en el FrontMath Benchmark antes de 2027?" con 90%, lo que significa que el mercado asigna una probabilidad de 90% a ese resultado. Estas probabilidades se actualizan en tiempo real a medida que los operadores compran y venden acciones. Vuelve con frecuencia o guarda esta página en marcadores.

Las reglas de resolución para "¿El modelo de IA obtiene un puntaje ≥ 90% en FrontierMath Benchmark antes de 2027?" definen exactamente qué debe ocurrir para que cada resultado sea declarado ganador, incluyendo las fuentes de datos oficiales utilizadas para determinar el resultado. Puedes revisar los criterios de resolución completos en la sección "Reglas" en esta página sobre los comentarios. Recomendamos leer las reglas cuidadosamente antes de operar, ya que especifican las condiciones exactas, casos especiales y fuentes.