Recent releases such as OpenAI’s GPT-5.6 Sol (July 2026) and Anthropic’s Claude Fable 5 and Mythos 5 have driven top coding Arena ELO scores and SWE-bench Verified results above 95 percent, reflecting strong trader consensus on continued rapid gains. Chinese labs including Moonshot (Kimi K3) and DeepSeek now compete directly on frontend and agentic coding leaderboards, narrowing the gap with proprietary frontier models. Key catalysts ahead include potential new model drops or fine-tunes from OpenAI, Anthropic, Google, and open-weight teams before year-end, alongside ongoing Arena updates that reward demonstrated code-generation and repository-level performance. Market-implied odds incorporate these capability jumps while accounting for typical release delays and benchmark saturation risks.
基于Polymarket数据的AI实验性摘要。这不是交易建议,也不影响该市场的结算方式。 · 更新于$185,292 交易量
1560
44%
1580
24%
1600
12%
$185,292 交易量
1560
44%
1580
24%
1600
12%
Results from the "Score" column under the "Text Arena | Coding" Leaderboard tab at https://arena.ai/leaderboard/text/coding-no-style-control with style control off will be used to resolve this market.
The resolution source for this market is the Chatbot Arena LLM Leaderboard found at arena.ai/leaderboard/text. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If permanently unavailable, this market will resolve to "No".
市场开放时间: Apr 2, 2026, 6:09 PM ET
Resolver
0x65070BE91...Results from the "Score" column under the "Text Arena | Coding" Leaderboard tab at https://arena.ai/leaderboard/text/coding-no-style-control with style control off will be used to resolve this market.
The resolution source for this market is the Chatbot Arena LLM Leaderboard found at arena.ai/leaderboard/text. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If permanently unavailable, this market will resolve to "No".
Resolver
0x65070BE91...Recent releases such as OpenAI’s GPT-5.6 Sol (July 2026) and Anthropic’s Claude Fable 5 and Mythos 5 have driven top coding Arena ELO scores and SWE-bench Verified results above 95 percent, reflecting strong trader consensus on continued rapid gains. Chinese labs including Moonshot (Kimi K3) and DeepSeek now compete directly on frontend and agentic coding leaderboards, narrowing the gap with proprietary frontier models. Key catalysts ahead include potential new model drops or fine-tunes from OpenAI, Anthropic, Google, and open-weight teams before year-end, alongside ongoing Arena updates that reward demonstrated code-generation and repository-level performance. Market-implied odds incorporate these capability jumps while accounting for typical release delays and benchmark saturation risks.
基于Polymarket数据的AI实验性摘要。这不是交易建议,也不影响该市场的结算方式。 · 更新于



警惕外部链接哦。
警惕外部链接哦。
常见问题