Traders see a tight contest for the leading AI agent by end of August, with Anthropic holding a slim 53% implied edge amid broad uncertainty reflected in multiple 49-50% outcomes. Recent model updates have highlighted Anthropic’s strengths in reliable tool use, long-horizon planning, and safety-aligned agent behavior, while OpenAI and others continue rapid iteration on multi-step reasoning and browser or code execution capabilities. Differentiating factors include demonstrated benchmark performance on complex tasks, developer ecosystem adoption, and regulatory scrutiny around autonomous systems. With roughly six weeks remaining, upcoming releases, capability demos, or third-party evaluations could quickly shift sentiment in this closely matched field.
Experimental AI-generated summary referencing Polymarket data. This is not trading advice and plays no role in how this market resolves. · UpdatedOpenAI 41%
Z.ai 41%
Moonshot 38%
Baidu 37.0%

OpenAI
41%

Z.ai
41%

Moonshot
38%

Baidu
37%

Alibaba
37%

33%

SpaceXAI
33%

MiniMax
33%

Xiaomi
33%

Microsoft
33%

Meta
33%

Nvidia
33%

ByteDance
33%

Mistral
33%

Amazon
32%

Meituan
28%

DeepSeek
23%

Anthropic
57%
OpenAI 41%
Z.ai 41%
Moonshot 38%
Baidu 37.0%

OpenAI
41%

Z.ai
41%

Moonshot
38%

Baidu
37%

Alibaba
37%

33%

SpaceXAI
33%

MiniMax
33%

Xiaomi
33%

Microsoft
33%

Meta
33%

Nvidia
33%

ByteDance
33%

Mistral
33%

Amazon
32%

Meituan
28%

DeepSeek
23%

Anthropic
57%
Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Models" will be used to resolve this market.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Agent Arena Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Market Opened: Jul 20, 2026, 7:06 PM ET
Resolution Source
https://arena.ai/leaderboard/agentResolver
0x69c47De9D...Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Models" will be used to resolve this market.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Agent Arena Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Resolution Source
https://arena.ai/leaderboard/agentResolver
0x69c47De9D...Traders see a tight contest for the leading AI agent by end of August, with Anthropic holding a slim 53% implied edge amid broad uncertainty reflected in multiple 49-50% outcomes. Recent model updates have highlighted Anthropic’s strengths in reliable tool use, long-horizon planning, and safety-aligned agent behavior, while OpenAI and others continue rapid iteration on multi-step reasoning and browser or code execution capabilities. Differentiating factors include demonstrated benchmark performance on complex tasks, developer ecosystem adoption, and regulatory scrutiny around autonomous systems. With roughly six weeks remaining, upcoming releases, capability demos, or third-party evaluations could quickly shift sentiment in this closely matched field.
Experimental AI-generated summary referencing Polymarket data. This is not trading advice and plays no role in how this market resolves. · Updated

Beware of external links.
Beware of external links.
Frequently Asked Questions