Moonshot AI’s July 2026 release of the 2.8-trillion-parameter Kimi K3 drove recent trader positioning on the Polymarket, with the model posting 46.9–56% on Humanity’s Last Exam depending on tool use and leaderboard protocol. This placed it competitively behind Anthropic’s Claude Fable 5.1 (59–65%) but ahead of most other open-weight entries. Subsequent Kimi iterations have shown incremental gains on agentic and reasoning subsets, yet the gap to closed frontier models persists amid rapid releases from OpenAI and Google. Traders are watching for any late-2026 Kimi updates or agent enhancements that could close the remaining distance before the December 31 resolution deadline, while noting that benchmark scores remain sensitive to exact evaluation conditions like tool access and reasoning effort.
Resumen experimental generado por IA con datos de Polymarket. Esto no es asesoramiento de trading y no influye en cómo se resuelve este mercado. · Actualizado$62,779 Vol.
45%+
85%
50%+
39%
55%+
19%
60% o más
11%
65% o más
7%
$62,779 Vol.
45%+
85%
50%+
39%
55%+
19%
60% o más
11%
65% o más
7%
For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Mercado abierto: Jul 23, 2026, 6:43 PM ET
Fuente de resolución
https://agi.safe.ai/Resolver
0x65070BE91...For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Fuente de resolución
https://agi.safe.ai/Resolver
0x65070BE91...Moonshot AI’s July 2026 release of the 2.8-trillion-parameter Kimi K3 drove recent trader positioning on the Polymarket, with the model posting 46.9–56% on Humanity’s Last Exam depending on tool use and leaderboard protocol. This placed it competitively behind Anthropic’s Claude Fable 5.1 (59–65%) but ahead of most other open-weight entries. Subsequent Kimi iterations have shown incremental gains on agentic and reasoning subsets, yet the gap to closed frontier models persists amid rapid releases from OpenAI and Google. Traders are watching for any late-2026 Kimi updates or agent enhancements that could close the remaining distance before the December 31 resolution deadline, while noting that benchmark scores remain sensitive to exact evaluation conditions like tool access and reasoning effort.
Resumen experimental generado por IA con datos de Polymarket. Esto no es asesoramiento de trading y no influye en cómo se resuelve este mercado. · Actualizado



Cuidado con los enlaces externos.
Cuidado con los enlaces externos.
Preguntas frecuentes