Anthropic's recent dominance in math-specific benchmarks and high-profile advances drives the 68.5% market-implied odds for its models leading the Text Arena Math category by end of October. Trader consensus reflects verified gains such as Claude Mythos Preview topping aggregated GSM8K, MATH, and AIME-style evaluations, alongside an unreleased research version advancing the Riemann hypothesis lower bound from 41.6% to 67.2% through extended agentic reasoning. These results outpace competitors including OpenAI and Google on frontier competition problems, though product timelines and benchmark saturation introduce ongoing uncertainty. No major releases from trailing labs have shifted positioning in the past month.
Eksperymentalne podsumowanie AI odwołujące się do danych Polymarket. To nie jest porada handlowa i nie ma wpływu na rozstrzyganie tego rynku. · ZaktualizowanoWhich company has the best Text Arena Math AI model end of October?
Anthropic 76%
OpenAI 13%
Google 11.9%
DeepSeek 3.7%

Anthropic
76%

OpenAI
13%

12%

DeepSeek
4%

Alibaba
2%

Tencent
2%

Xiaomi
2%

SpaceXAI
1%

Moonshot
1%

Meta
1%

StepFun
1%

Meituan
1%

Amazon
1%

Z.ai
1%

Nvidia
1%

MiniMax
1%

ByteDance
1%

Mistral
1%

Microsoft
1%

Baidu
<1%
Anthropic 76%
OpenAI 13%
Google 11.9%
DeepSeek 3.7%

Anthropic
76%

OpenAI
13%

12%

DeepSeek
4%

Alibaba
2%

Tencent
2%

Xiaomi
2%

SpaceXAI
1%

Moonshot
1%

Meta
1%

StepFun
1%

Meituan
1%

Amazon
1%

Z.ai
1%

Nvidia
1%

MiniMax
1%

ByteDance
1%

Mistral
1%

Microsoft
1%

Baidu
<1%
Results from the "Rank" column under the "Text Arena | Math" Leaderboard tab at https://arena.ai/leaderboard/text/math-no-style-control with style control off (Adjustments: None) and filtered for "Models" will be used to resolve this market.
Note: Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by their Arena score, including any underlying, unrounded, granular values reflected in the data below the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact arena score, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Text Arena (Math). If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Rynek otwarty: Aug 12, 2026, 7:52 PM ET
Resolver
0x69c47De9D...Results from the "Rank" column under the "Text Arena | Math" Leaderboard tab at https://arena.ai/leaderboard/text/math-no-style-control with style control off (Adjustments: None) and filtered for "Models" will be used to resolve this market.
Note: Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by their Arena score, including any underlying, unrounded, granular values reflected in the data below the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact arena score, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Text Arena (Math). If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Resolver
0x69c47De9D...Anthropic's recent dominance in math-specific benchmarks and high-profile advances drives the 68.5% market-implied odds for its models leading the Text Arena Math category by end of October. Trader consensus reflects verified gains such as Claude Mythos Preview topping aggregated GSM8K, MATH, and AIME-style evaluations, alongside an unreleased research version advancing the Riemann hypothesis lower bound from 41.6% to 67.2% through extended agentic reasoning. These results outpace competitors including OpenAI and Google on frontier competition problems, though product timelines and benchmark saturation introduce ongoing uncertainty. No major releases from trailing labs have shifted positioning in the past month.
Eksperymentalne podsumowanie AI odwołujące się do danych Polymarket. To nie jest porada handlowa i nie ma wpływu na rozstrzyganie tego rynku. · Zaktualizowano
Uważaj na linki zewnętrzne.
Uważaj na linki zewnętrzne.
Często zadawane pytania