Google’s recent Gemini 4 Argon release has driven strong gains on the Text Arena Math leaderboard, where its high-effort variant currently tops crowd-sourced math evaluations with scores above 1530, ahead of Anthropic’s Claude Opus 5 and Fable 5 variants clustered just behind. Traders assign Anthropic a slim 49% edge for end-December leadership, reflecting its consistent math reasoning track record, while Google sits at 22% amid expectations of further iteration. Lower probabilities for OpenAI, Meta, and Chinese labs like Z.ai, Moonshot, and DeepSeek capture the fragmented competitive landscape, where specialized advances in multi-step problem solving, tool integration, and benchmark-specific tuning can rapidly alter arena standings before resolution. New model drops and arena vote surges remain the key swing factors in this tight contest.
Polymarketデータを参照したAI生成の実験的な要約。これは取引アドバイスではなく、このマーケットの解決方法には一切関係ありません。 · 更新日View resolved






















外部リンクに注意してください。
外部リンクに注意してください。
よくある質問