OpenAI’s recent launch of GPT-6 Astra, which quickly claimed the top spot on the Code Arena WebDev leaderboard with the highest arena score around 1,797–1,800 as of early September, drives the 76.7% market-implied odds for the company. The model’s edge in agentic frontend coding tasks, multi-step reasoning, and tool use reflects strong performance on real-world web development benchmarks that emphasize iterative HTML, React, and UI generation workflows. Anthropic’s Claude Fable 5.1 series sits a clear second at roughly 23% odds, trailing by about 35 points but still competitive among proprietary large language models. Most other labs remain far behind on the latest community-voted rankings, with limited recent releases or benchmark gains shifting sentiment before the end-of-September resolution.
Експериментальне резюме, згенероване ШІ з посиланням на дані Polymarket. Це не торгова порада і не впливає на вирішення цього ринку. · ОновленоWhich company has the best Code Arena WebDev AI model end of September?
OpenAI 76.7%
Anthropic 23%
Google <1%
Alibaba <1%
$145,677 Обс.
$145,677 Обс.

OpenAI
77%

Anthropic
23%

<1%

Alibaba
<1%

Moonshot
<1%

Z.ai
<1%

SpaceXAI
<1%

Meta
<1%

ByteDance
<1%

MiniMax
<1%

Xiaomi
<1%

DeepSeek
<1%

Thinky
<1%

Tencent
<1%

Poolside
<1%

Mistral
<1%
OpenAI 76.7%
Anthropic 23%
Google <1%
Alibaba <1%
$145,677 Обс.
$145,677 Обс.

OpenAI
77%

Anthropic
23%

<1%

Alibaba
<1%

Moonshot
<1%

Z.ai
<1%

SpaceXAI
<1%

Meta
<1%

ByteDance
<1%

MiniMax
<1%

Xiaomi
<1%

DeepSeek
<1%

Thinky
<1%

Tencent
<1%

Poolside
<1%

Mistral
<1%
Results from the "Rank" column under the "Code Arena | WebDev" Leaderboard tab at https://arena.ai/leaderboard/code/webdev (Overall) filtered for "Models" will be used to resolve this market.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Code Arena | WebDev Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Ринок відкрито: Jul 20, 2026, 7:23 PM ET
Джерело вирішення
https://arena.ai/leaderboard/code/webdevВирішувач
0x69c47De9D...Results from the "Rank" column under the "Code Arena | WebDev" Leaderboard tab at https://arena.ai/leaderboard/code/webdev (Overall) filtered for "Models" will be used to resolve this market.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Code Arena | WebDev Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Джерело вирішення
https://arena.ai/leaderboard/code/webdevВирішувач
0x69c47De9D...OpenAI’s recent launch of GPT-6 Astra, which quickly claimed the top spot on the Code Arena WebDev leaderboard with the highest arena score around 1,797–1,800 as of early September, drives the 76.7% market-implied odds for the company. The model’s edge in agentic frontend coding tasks, multi-step reasoning, and tool use reflects strong performance on real-world web development benchmarks that emphasize iterative HTML, React, and UI generation workflows. Anthropic’s Claude Fable 5.1 series sits a clear second at roughly 23% odds, trailing by about 35 points but still competitive among proprietary large language models. Most other labs remain far behind on the latest community-voted rankings, with limited recent releases or benchmark gains shifting sentiment before the end-of-September resolution.
Експериментальне резюме, згенероване ШІ з посиланням на дані Polymarket. Це не торгова порада і не впливає на вирішення цього ринку. · Оновлено
Обережно з зовнішніми посиланнями.
Обережно з зовнішніми посиланнями.
Часті запитання