Anthropic’s release of detailed R&D automation metrics in mid-September, showing Claude leading 26% of internal AI research and development work while collaborating on over 90%, has anchored trader consensus at 92% for the company holding the best AI agent by end of September. Recent Claude models such as Fable 5.1 and Opus 5.1 continue to top independent agentic benchmarks measuring tool use, long-horizon task completion, and error recovery, reinforced by practical releases like the open-source Commerce Agents blueprint and Project Fetch demonstrations of physical tool operation. Traders view these verified capability and safety transparency steps as decisive in a market where agent performance hinges on reliable multi-step execution rather than raw model scale. A late-month OpenAI model update or undisclosed benchmark reversal could still shift sentiment before resolution.
Polymarketデータを参照したAI生成の実験的な要約。これは取引アドバイスではなく、このマーケットの解決方法には一切関係ありません。 · 更新日Anthropic 92%
OpenAI 6.2%
マイクロソフト <1%
Google <1%
$223,979 Vol.
$223,979 Vol.

Anthropic
92%

OpenAI
6%

マイクロソフト
<1%

<1%

Z.ai
<1%

スペースXAI
<1%

Meta
<1%

アリババ
<1%

Moonshot
<1%

DeepSeek
<1%

Xiaomi
<1%

MiniMax
<1%

Nvidia
<1%

バイトダンス
<1%

バイドゥ
<1%

アマゾン
<1%

Mistral
<1%

美団
<1%
Anthropic 92%
OpenAI 6.2%
マイクロソフト <1%
Google <1%
$223,979 Vol.
$223,979 Vol.

Anthropic
92%

OpenAI
6%

マイクロソフト
<1%

<1%

Z.ai
<1%

スペースXAI
<1%

Meta
<1%

アリババ
<1%

Moonshot
<1%

DeepSeek
<1%

Xiaomi
<1%

MiniMax
<1%

Nvidia
<1%

バイトダンス
<1%

バイドゥ
<1%

アマゾン
<1%

Mistral
<1%

美団
<1%
Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Models" will be used to resolve this market.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the Agent Arena Leaderboard found at https://arena.ai/leaderboard/agent. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve based on another resolution source.
マーケット開始日: Jul 20, 2026, 7:06 PM ET
リゾルバー
0x69c47De9D...Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Models" will be used to resolve this market.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the Agent Arena Leaderboard found at https://arena.ai/leaderboard/agent. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve based on another resolution source.
リゾルバー
0x69c47De9D...Anthropic’s release of detailed R&D automation metrics in mid-September, showing Claude leading 26% of internal AI research and development work while collaborating on over 90%, has anchored trader consensus at 92% for the company holding the best AI agent by end of September. Recent Claude models such as Fable 5.1 and Opus 5.1 continue to top independent agentic benchmarks measuring tool use, long-horizon task completion, and error recovery, reinforced by practical releases like the open-source Commerce Agents blueprint and Project Fetch demonstrations of physical tool operation. Traders view these verified capability and safety transparency steps as decisive in a market where agent performance hinges on reliable multi-step execution rather than raw model scale. A late-month OpenAI model update or undisclosed benchmark reversal could still shift sentiment before resolution.
Polymarketデータを参照したAI生成の実験的な要約。これは取引アドバイスではなく、このマーケットの解決方法には一切関係ありません。 · 更新日
外部リンクに注意してください。
外部リンクに注意してください。
よくある質問