Anthropic's Claude models hold a commanding 78.5% implied probability in the Code Arena WebDev market due to sustained leadership in agentic coding benchmarks and real-world frontend workflows. Recent releases, including Opus and Sonnet variants, have demonstrated top scores on SWE-bench Verified and similar evaluations measuring multi-step reasoning, tool use, and long-running web app generation tasks. This edge stems from iterative improvements in context handling and reliability that align closely with Code Arena's emphasis on practical front-end development. Other labs like OpenAI, Meta, and DeepSeek trail significantly, reflecting narrower gaps in direct comparisons of web-specific capabilities. With the October resolution still months away, traders are watching for new model drops or leaderboard updates that could shift the current consensus.
Resumen experimental generado por IA con datos de Polymarket. Esto no es asesoramiento de trading y no influye en cómo se resuelve este mercado. · ActualizadoAnthropic 73%
OpenAI 10%
Meta 9.5%
Thinky 5.0%
$13,308 Vol.
$13,308 Vol.

Anthropic
73%

OpenAI
10%

Meta
10%

Thinky
5%

SpaceXAI
5%

Alibaba
3%

3%

Tencent
3%

Xiaomi
3%

Poolside
3%

Moonshot
2%

Z.ai
1%

DeepSeek
1%

Mistral
1%

ByteDance
<1%

MiniMax
<1%
Anthropic 73%
OpenAI 10%
Meta 9.5%
Thinky 5.0%
$13,308 Vol.
$13,308 Vol.

Anthropic
73%

OpenAI
10%

Meta
10%

Thinky
5%

SpaceXAI
5%

Alibaba
3%

3%

Tencent
3%

Xiaomi
3%

Poolside
3%

Moonshot
2%

Z.ai
1%

DeepSeek
1%

Mistral
1%

ByteDance
<1%

MiniMax
<1%
Results from the "Rank" column under the "Code Arena | WebDev" Leaderboard tab at https://arena.ai/leaderboard/code/webdev (Overall) filtered for "Models" will be used to resolve this market.
Note: Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Code Arena | WebDev Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Mercado abierto: Aug 12, 2026, 7:56 PM ET
Resolver
0x69c47De9D...Results from the "Rank" column under the "Code Arena | WebDev" Leaderboard tab at https://arena.ai/leaderboard/code/webdev (Overall) filtered for "Models" will be used to resolve this market.
Note: Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Code Arena | WebDev Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Resolver
0x69c47De9D...Anthropic's Claude models hold a commanding 78.5% implied probability in the Code Arena WebDev market due to sustained leadership in agentic coding benchmarks and real-world frontend workflows. Recent releases, including Opus and Sonnet variants, have demonstrated top scores on SWE-bench Verified and similar evaluations measuring multi-step reasoning, tool use, and long-running web app generation tasks. This edge stems from iterative improvements in context handling and reliability that align closely with Code Arena's emphasis on practical front-end development. Other labs like OpenAI, Meta, and DeepSeek trail significantly, reflecting narrower gaps in direct comparisons of web-specific capabilities. With the October resolution still months away, traders are watching for new model drops or leaderboard updates that could shift the current consensus.
Resumen experimental generado por IA con datos de Polymarket. Esto no es asesoramiento de trading y no influye en cómo se resuelve este mercado. · Actualizado
Cuidado con los enlaces externos.
Cuidado con los enlaces externos.
Preguntas frecuentes