OpenAI’s recent launch of GPT-6 Astra on September 3 propelled the model to the top of the Code Arena WebDev leaderboard with a score of roughly 1,797, edging out Anthropic’s Claude Fable 5.1 by 35 points on the community-driven Elo-style ranking of agentic coding and frontend web development tasks. The crowdsourced benchmark, which aggregates hundreds of thousands of votes on multi-step reasoning, tool use, and iterative workflows, reflects OpenAI’s gains in practical web dev capabilities over prior GPT-5.6 variants. Anthropic maintains a credible second-place position with its latest Claude release, while Google’s Gemini 3.7 Flash and other labs trail further behind on the same metrics. With the evaluation window closing at month-end, traders are pricing in limited time for competitors to close the gap through new model updates or benchmark shifts.
Tóm tắt AI thử nghiệm tham chiếu dữ liệu Polymarket. Đây không phải tư vấn giao dịch và không ảnh hưởng đến cách thị trường này được giải quyết. · Cập nhậtOpenAI 76.4%
Anthropic 25%
Google <1%
Alibaba <1%
$143,938 KL.
$143,938 KL.

OpenAI
76%

Anthropic
25%

<1%

Alibaba
<1%

Moonshot
<1%

Z.ai
<1%

SpaceXAI
<1%

Meta
<1%

ByteDance
<1%

MiniMax
<1%

Xiaomi
<1%

DeepSeek
<1%

Thinky
<1%

Tencent
<1%

Poolside
<1%

Mistral
<1%
OpenAI 76.4%
Anthropic 25%
Google <1%
Alibaba <1%
$143,938 KL.
$143,938 KL.

OpenAI
76%

Anthropic
25%

<1%

Alibaba
<1%

Moonshot
<1%

Z.ai
<1%

SpaceXAI
<1%

Meta
<1%

ByteDance
<1%

MiniMax
<1%

Xiaomi
<1%

DeepSeek
<1%

Thinky
<1%

Tencent
<1%

Poolside
<1%

Mistral
<1%
Results from the "Rank" column under the "Code Arena | WebDev" Leaderboard tab at https://arena.ai/leaderboard/code/webdev (Overall) filtered for "Models" will be used to resolve this market.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Code Arena | WebDev Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Thị trường mở: Jul 20, 2026, 7:23 PM ET
Nguồn giải quyết
https://arena.ai/leaderboard/code/webdevNgười giải quyết
0x69c47De9D...Results from the "Rank" column under the "Code Arena | WebDev" Leaderboard tab at https://arena.ai/leaderboard/code/webdev (Overall) filtered for "Models" will be used to resolve this market.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Code Arena | WebDev Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Nguồn giải quyết
https://arena.ai/leaderboard/code/webdevNgười giải quyết
0x69c47De9D...OpenAI’s recent launch of GPT-6 Astra on September 3 propelled the model to the top of the Code Arena WebDev leaderboard with a score of roughly 1,797, edging out Anthropic’s Claude Fable 5.1 by 35 points on the community-driven Elo-style ranking of agentic coding and frontend web development tasks. The crowdsourced benchmark, which aggregates hundreds of thousands of votes on multi-step reasoning, tool use, and iterative workflows, reflects OpenAI’s gains in practical web dev capabilities over prior GPT-5.6 variants. Anthropic maintains a credible second-place position with its latest Claude release, while Google’s Gemini 3.7 Flash and other labs trail further behind on the same metrics. With the evaluation window closing at month-end, traders are pricing in limited time for competitors to close the gap through new model updates or benchmark shifts.
Tóm tắt AI thử nghiệm tham chiếu dữ liệu Polymarket. Đây không phải tư vấn giao dịch và không ảnh hưởng đến cách thị trường này được giải quyết. · Cập nhật
Cẩn thận với liên kết bên ngoài.
Cẩn thận với liên kết bên ngoài.
Câu hỏi thường gặp