Anthropic’s commanding 96.3% market-implied probability stems from its Claude models’ leading performance on independent agentic benchmarks such as SWE-bench Verified (96% for Opus 5) and strong results across computer-use, browser, and long-horizon tool-use evaluations. The August 19 general availability of production-grade computer and browser toolsets, the Agent Skills API, and recent Claude Tag updates for proactive multi-user workflows have reinforced developer and enterprise adoption. With only days remaining until the August 31 resolution on the arena.ai Agent Arena leaderboard, no rival has released comparable capabilities. A last-minute benchmark surge or model launch from OpenAI or Google could theoretically narrow the gap, though current timelines make such shifts improbable.
Experimental AI-generated summary referencing Polymarket data. This is not trading advice and plays no role in how this market resolves. · UpdatedAnthropic 96.2%
OpenAI 1.9%
Alibaba <1%
Z.ai <1%
$119,938 Vol.
$119,938 Vol.

Anthropic
96%

OpenAI
2%

Alibaba
1%

Z.ai
1%

Meta
1%

SpaceXAI
<1%

MiniMax
<1%

<1%

Moonshot
<1%

Microsoft
<1%

Baidu
<1%

Xiaomi
<1%

Amazon
<1%

DeepSeek
<1%

Nvidia
<1%

ByteDance
<1%

Mistral
<1%

Meituan
<1%
Anthropic 96.2%
OpenAI 1.9%
Alibaba <1%
Z.ai <1%
$119,938 Vol.
$119,938 Vol.

Anthropic
96%

OpenAI
2%

Alibaba
1%

Z.ai
1%

Meta
1%

SpaceXAI
<1%

MiniMax
<1%

<1%

Moonshot
<1%

Microsoft
<1%

Baidu
<1%

Xiaomi
<1%

Amazon
<1%

DeepSeek
<1%

Nvidia
<1%

ByteDance
<1%

Mistral
<1%

Meituan
<1%
Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Models" will be used to resolve this market.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Agent Arena Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Market Opened: Jul 20, 2026, 7:06 PM ET
Resolution Source
https://arena.ai/leaderboard/agentResolver
0x69c47De9D...Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Models" will be used to resolve this market.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Agent Arena Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Resolution Source
https://arena.ai/leaderboard/agentResolver
0x69c47De9D...Anthropic’s commanding 96.3% market-implied probability stems from its Claude models’ leading performance on independent agentic benchmarks such as SWE-bench Verified (96% for Opus 5) and strong results across computer-use, browser, and long-horizon tool-use evaluations. The August 19 general availability of production-grade computer and browser toolsets, the Agent Skills API, and recent Claude Tag updates for proactive multi-user workflows have reinforced developer and enterprise adoption. With only days remaining until the August 31 resolution on the arena.ai Agent Arena leaderboard, no rival has released comparable capabilities. A last-minute benchmark surge or model launch from OpenAI or Google could theoretically narrow the gap, though current timelines make such shifts improbable.
Experimental AI-generated summary referencing Polymarket data. This is not trading advice and plays no role in how this market resolves. · Updated
Beware of external links.
Beware of external links.
Frequently Asked Questions