Skip to main content

Para operar en EE. UU., ve a polymarket.us

icon for Next Grok Model (4.6+): Humanity’s Last Exam Debut?

Next Grok Model (4.6+): Humanity’s Last Exam Debut?

icon for Next Grok Model (4.6+): Humanity’s Last Exam Debut?

Next Grok Model (4.6+): Humanity’s Last Exam Debut?

$29,129 Vol.

31 dic 2026
Polymarket

$29,129 Vol.

Polymarket

35%+

$4,090 Vol.

Yes

40%+

$9,105 Vol.

No

45%+

$9,706 Vol.

No

50%+

$3,907 Vol.

No

55%+

$2,321 Vol.

No

This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No". If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results. A qualifying SpaceXAI model must have “Grok” in its displayed model name and be designated as version 4.6 or higher, regardless of capitalization or surrounding prefixes, suffixes, dates, or descriptors. For example, grok-4.6-high, grok-4.7-thinking, grok-5, or similar would qualify. Models whose displayed name does not include “Grok,” or which retain a version designation below 4.6, such as grok-4.1-expert, grok-4.3-heavy, grok-4.5-GA will not qualify. The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere. If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered. A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market. The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".**Grok 4.6, xAI’s 1.5-trillion-parameter successor to Grok 4.5, shipped in early August 2026 with post-training upgrades targeting coding, agentic workflows, and engineering tasks.** Prior Grok 4 variants posted competitive HLE results, including up to 50% on multi-agent “Heavy” configurations, though public leaderboards remain led by Anthropic’s Claude Fable 5.1 and Opus 5 at 64.5–65%. The market reflects trader focus on whether the latest Grok release will promptly appear on the official HLE leaderboard at frontier-comparable accuracy, given xAI’s pattern of benchmark emphasis and the benchmark’s rapid score climb from single-digit percentages in early 2025 to the mid-40s–60s% range today. Key near-term catalysts include any verified HLE submission, a potential Grok 4.7 follow-up, and ongoing competition in reasoning and expert-knowledge evaluations.

This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No".

If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results.

A qualifying SpaceXAI model must have “Grok” in its displayed model name and be designated as version 4.6 or higher, regardless of capitalization or surrounding prefixes, suffixes, dates, or descriptors. For example, grok-4.6-high, grok-4.7-thinking, grok-5, or similar would qualify. Models whose displayed name does not include “Grok,” or which retain a version designation below 4.6, such as grok-4.1-expert, grok-4.3-heavy, grok-4.5-GA will not qualify.

The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere.

If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered.

A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market.

The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".
This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No". If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results. A qualifying SpaceXAI model must have “Grok” in its displayed model name and be designated as version 4.6 or higher, regardless of capitalization or surrounding prefixes, suffixes, dates, or descriptors. For example, grok-4.6-high, grok-4.7-thinking, grok-5, or similar would qualify. Models whose displayed name does not include “Grok,” or which retain a version designation below 4.6, such as grok-4.1-expert, grok-4.3-heavy, grok-4.5-GA will not qualify. The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere. If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered. A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market. The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".
Volumen
$29,129
Fecha de finalización
31 dic 2026
Mercado abierto
Aug 10, 2026, 6:25 PM ET

Fuente de resolución

https://agi.safe.ai/

Resultado propuesto: Yes

Sin disputa

Resultado final: Yes

This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No". If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results. A qualifying SpaceXAI model must have “Grok” in its displayed model name and be designated as version 4.6 or higher, regardless of capitalization or surrounding prefixes, suffixes, dates, or descriptors. For example, grok-4.6-high, grok-4.7-thinking, grok-5, or similar would qualify. Models whose displayed name does not include “Grok,” or which retain a version designation below 4.6, such as grok-4.1-expert, grok-4.3-heavy, grok-4.5-GA will not qualify. The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere. If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered. A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market. The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".**Grok 4.6, xAI’s 1.5-trillion-parameter successor to Grok 4.5, shipped in early August 2026 with post-training upgrades targeting coding, agentic workflows, and engineering tasks.** Prior Grok 4 variants posted competitive HLE results, including up to 50% on multi-agent “Heavy” configurations, though public leaderboards remain led by Anthropic’s Claude Fable 5.1 and Opus 5 at 64.5–65%. The market reflects trader focus on whether the latest Grok release will promptly appear on the official HLE leaderboard at frontier-comparable accuracy, given xAI’s pattern of benchmark emphasis and the benchmark’s rapid score climb from single-digit percentages in early 2025 to the mid-40s–60s% range today. Key near-term catalysts include any verified HLE submission, a potential Grok 4.7 follow-up, and ongoing competition in reasoning and expert-knowledge evaluations.

This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No".

If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results.

A qualifying SpaceXAI model must have “Grok” in its displayed model name and be designated as version 4.6 or higher, regardless of capitalization or surrounding prefixes, suffixes, dates, or descriptors. For example, grok-4.6-high, grok-4.7-thinking, grok-5, or similar would qualify. Models whose displayed name does not include “Grok,” or which retain a version designation below 4.6, such as grok-4.1-expert, grok-4.3-heavy, grok-4.5-GA will not qualify.

The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere.

If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered.

A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market.

The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".
This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No". If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results. A qualifying SpaceXAI model must have “Grok” in its displayed model name and be designated as version 4.6 or higher, regardless of capitalization or surrounding prefixes, suffixes, dates, or descriptors. For example, grok-4.6-high, grok-4.7-thinking, grok-5, or similar would qualify. Models whose displayed name does not include “Grok,” or which retain a version designation below 4.6, such as grok-4.1-expert, grok-4.3-heavy, grok-4.5-GA will not qualify. The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere. If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered. A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market. The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".
Volumen
$29,129
Fecha de finalización
31 dic 2026
Mercado abierto
Aug 10, 2026, 6:25 PM ET

Fuente de resolución

https://agi.safe.ai/

Resultado propuesto: Yes

Sin disputa

Resultado final: Yes

Cuidado con los enlaces externos.

Preguntas frecuentes

"Next Grok Model (4.6+): Humanity’s Last Exam Debut?" es un mercado de predicción en Polymarket con 5 resultados posibles donde los operadores compran y venden acciones según lo que creen que sucederá. El resultado líder actual es "35%+" con 100%, seguido de "40%+" con 0%. Los precios reflejan probabilidades en tiempo real de la comunidad. Por ejemplo, una acción cotizada a 100¢ implica que el mercado colectivamente asigna una probabilidad de 100% a ese resultado. Estas probabilidades cambian continuamente a medida que los operadores reaccionan a nuevos desarrollos. Las acciones del resultado correcto son canjeables por $1 cada una tras la resolución del mercado.

A día de hoy, "Next Grok Model (4.6+): Humanity’s Last Exam Debut?" ha generado $29.1K en volumen total de trading desde que el mercado se lanzó el Aug 11, 2026. Este nivel de actividad refleja un fuerte compromiso de la comunidad de Polymarket y ayuda a garantizar que las probabilidades actuales estén respaldadas por un amplio grupo de participantes del mercado. Puedes seguir los movimientos de precios en vivo y operar en cualquier resultado directamente en esta página.

Para operar en "Next Grok Model (4.6+): Humanity’s Last Exam Debut?", explora los 5 resultados disponibles en esta página. Cada resultado muestra un precio actual que representa la probabilidad implícita del mercado. Para tomar una posición, selecciona el resultado que consideres más probable, elige "Sí" para operar a favor o "No" para operar en contra, introduce tu cantidad y haz clic en "Operar". Si tu resultado elegido es correcto cuando el mercado se resuelve, tus acciones de "Sí" pagan $1 cada una. Si es incorrecto, pagan $0. También puedes vender tus acciones en cualquier momento antes de la resolución.

El favorito actual para "Next Grok Model (4.6+): Humanity’s Last Exam Debut?" es "35%+" con 100%, lo que significa que el mercado asigna una probabilidad de 100% a ese resultado. El siguiente resultado más cercano es "40%+" con 0%. Estas probabilidades se actualizan en tiempo real a medida que los operadores compran y venden acciones. Vuelve con frecuencia o guarda esta página en marcadores.

Las reglas de resolución para "Next Grok Model (4.6+): Humanity’s Last Exam Debut?" definen exactamente qué debe ocurrir para que cada resultado sea declarado ganador, incluyendo las fuentes de datos oficiales utilizadas para determinar el resultado. Puedes revisar los criterios de resolución completos en la sección "Reglas" en esta página sobre los comentarios. Recomendamos leer las reglas cuidadosamente antes de operar, ya que especifican las condiciones exactas, casos especiales y fuentes.