Recent September 2026 model releases have fragmented the AI agent landscape, with no single leader ahead of the November deadline. Anthropic’s Claude Fable 5.1 and Mythos 5.1 lead long-horizon autonomy and Terminal-Bench scores, while OpenAI’s GPT-6 Astra excels on ARC-AGI, ExploitBench, and computer-use tasks; its new Agents API further boosts practical deployment via managed multi-agent workflows and sandboxes. Open-weight contenders from Z.ai (GLM), DeepSeek, and Alibaba (Qwen) post competitive agentic benchmarks at lower cost, sustaining trader parity. Safety coordination among OpenAI, Anthropic, and Google adds regulatory uncertainty, keeping implied probabilities closely matched across labs.
Experimental AI-generated summary referencing Polymarket data. This is not trading advice and plays no role in how this market resolves. · UpdatedNvidia 51%
SpaceXAI 47%
DeepSeek 47%
Moonshot 46%

Nvidia
51%

SpaceXAI
47%

DeepSeek
47%

Moonshot
46%

Baidu
46%

Meituan
46%

Microsoft
46%

Anthropic
61%

Mistral
41%

Alibaba
28%

OpenAI
26%

Z.ai
26%

MiniMax
25%

23%

Meta
14%

Amazon
14%

ByteDance
14%

Xiaomi
25%

Tencent
8%
Nvidia 51%
SpaceXAI 47%
DeepSeek 47%
Moonshot 46%

Nvidia
51%

SpaceXAI
47%

DeepSeek
47%

Moonshot
46%

Baidu
46%

Meituan
46%

Microsoft
46%

Anthropic
61%

Mistral
41%

Alibaba
28%

OpenAI
26%

Z.ai
26%

MiniMax
25%

23%

Meta
14%

Amazon
14%

ByteDance
14%

Xiaomi
25%

Tencent
8%
Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Models" will be used to resolve this market.
Note: Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the Agent Arena Leaderboard found at https://arena.ai/leaderboard/agent. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve based on another resolution source.
Market Opened: Sep 17, 2026, 8:02 PM ET
Resolver
0x69c47De9D...Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Models" will be used to resolve this market.
Note: Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the Agent Arena Leaderboard found at https://arena.ai/leaderboard/agent. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve based on another resolution source.
Resolver
0x69c47De9D...Recent September 2026 model releases have fragmented the AI agent landscape, with no single leader ahead of the November deadline. Anthropic’s Claude Fable 5.1 and Mythos 5.1 lead long-horizon autonomy and Terminal-Bench scores, while OpenAI’s GPT-6 Astra excels on ARC-AGI, ExploitBench, and computer-use tasks; its new Agents API further boosts practical deployment via managed multi-agent workflows and sandboxes. Open-weight contenders from Z.ai (GLM), DeepSeek, and Alibaba (Qwen) post competitive agentic benchmarks at lower cost, sustaining trader parity. Safety coordination among OpenAI, Anthropic, and Google adds regulatory uncertainty, keeping implied probabilities closely matched across labs.
Experimental AI-generated summary referencing Polymarket data. This is not trading advice and plays no role in how this market resolves. · Updated
Beware of external links.
Beware of external links.
Frequently Asked Questions