Anthropic holds the strongest market-implied odds for topping LiveBench coding by end of October, driven by consistent leadership in recent Claude model releases on coding-specific tasks and agentic benchmarks. The Fable 5.1 and related Opus variants have posted top scores in code generation, completion, and real-world software engineering evaluations, outpacing OpenAI's GPT-5.6 and GPT-6 series in several independent runs despite OpenAI's overall benchmark strength. Traders appear to weigh Anthropic's specialized post-training and demonstrated edge in practical coding workflows more heavily than broader model capabilities. With roughly six weeks remaining, any new frontier releases or LiveBench question refreshes could shift sentiment, though the current gap favors Anthropic's positioning in this subcategory.
Polymarket ডেটা রেফারেন্স করে পরীক্ষামূলক AI-জেনারেটেড সারাংশ। এটি ট্রেডিং পরামর্শ নয় এবং এই মার্কেট কীভাবে রেজলভ হয় তাতে কোনো ভূমিকা রাখে না। · আপডেটেডAnthropic 79%
OpenAI 6.6%
SpaceXAI <1%
Google <1%
$21,757 Vol.
$21,757 Vol.

Anthropic
79%

OpenAI
7%

SpaceXAI
1%

1%

DeepSeek
1%

Mistral
<1%

Nvidia
<1%

MiniMax
<1%

Meta
<1%

Alibaba
<1%

Moonshot
<1%

Baidu
<1%

Z.ai
<1%

Xiaomi
<1%

Amazon
<1%

ByteDance
<1%

Thinky
<1%

Meituan
<1%

Tencent
<1%

StepFun
<1%

Microsoft
<1%
Anthropic 79%
OpenAI 6.6%
SpaceXAI <1%
Google <1%
$21,757 Vol.
$21,757 Vol.

Anthropic
79%

OpenAI
7%

SpaceXAI
1%

1%

DeepSeek
1%

Mistral
<1%

Nvidia
<1%

MiniMax
<1%

Meta
<1%

Alibaba
<1%

Moonshot
<1%

Baidu
<1%

Z.ai
<1%

Xiaomi
<1%

Amazon
<1%

ByteDance
<1%

Thinky
<1%

Meituan
<1%

Tencent
<1%

StepFun
<1%

Microsoft
<1%
Results from the “Coding” column of the leaderboard at https://livebench.ai/#/?cats=Coding, with the latest available LiveBench release selected and the category set to “Coding,” will be used to resolve this market.
Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
মার্কেট ওপেন হয়েছে: Aug 12, 2026, 8:05 PM ET
রেজলভার
0x69c47De9D...Results from the “Coding” column of the leaderboard at https://livebench.ai/#/?cats=Coding, with the latest available LiveBench release selected and the category set to “Coding,” will be used to resolve this market.
Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
রেজলভার
0x69c47De9D...Anthropic holds the strongest market-implied odds for topping LiveBench coding by end of October, driven by consistent leadership in recent Claude model releases on coding-specific tasks and agentic benchmarks. The Fable 5.1 and related Opus variants have posted top scores in code generation, completion, and real-world software engineering evaluations, outpacing OpenAI's GPT-5.6 and GPT-6 series in several independent runs despite OpenAI's overall benchmark strength. Traders appear to weigh Anthropic's specialized post-training and demonstrated edge in practical coding workflows more heavily than broader model capabilities. With roughly six weeks remaining, any new frontier releases or LiveBench question refreshes could shift sentiment, though the current gap favors Anthropic's positioning in this subcategory.
Polymarket ডেটা রেফারেন্স করে পরীক্ষামূলক AI-জেনারেটেড সারাংশ। এটি ট্রেডিং পরামর্শ নয় এবং এই মার্কেট কীভাবে রেজলভ হয় তাতে কোনো ভূমিকা রাখে না। · আপডেটেড
বাহ্যিক লিংক থেকে সাবধান।
বাহ্যিক লিংক থেকে সাবধান।
সচরাচর জিজ্ঞাসা