Anthropic’s rapid 2026 releases of Claude Opus 5 and Opus 5.5 have driven strong trader focus on HLE performance, with the latter achieving a leading 67.7% on tool-augmented evaluations and 61.4% on verified no-tools runs shortly after its September 22 launch. These models, alongside variants like Fable 5.1 and Mythos 5, consistently top leaderboards on the 2,500-question expert benchmark covering advanced math, science, and humanities, outpacing OpenAI and Google entries by several points amid a landscape where scores remain well below saturation. Competitive positioning favors Anthropic’s emphasis on agentic reasoning and long-context capabilities, while upcoming catalysts include potential further Opus iterations or benchmark updates that could shift implied probabilities based on verified capability jumps.
Eksperymentalne podsumowanie AI odwołujące się do danych Polymarket. To nie jest porada handlowa i nie ma wpływu na rozstrzyganie tego rynku. · ZaktualizowanoView resolved

Uważaj na linki zewnętrzne.
Uważaj na linki zewnętrzne.
Często zadawane pytania