Recent September 2026 model releases have fragmented the AI agent landscape, with no single leader ahead of the November deadline. Anthropic’s Claude Fable 5.1 and Mythos 5.1 lead long-horizon autonomy and Terminal-Bench scores, while OpenAI’s GPT-6 Astra excels on ARC-AGI, ExploitBench, and computer-use tasks; its new Agents API further boosts practical deployment via managed multi-agent workflows and sandboxes. Open-weight contenders from Z.ai (GLM), DeepSeek, and Alibaba (Qwen) post competitive agentic benchmarks at lower cost, sustaining trader parity. Safety coordination among OpenAI, Anthropic, and Google adds regulatory uncertainty, keeping implied probabilities closely matched across labs.
Experimentelle KI-generierte Zusammenfassung mit Polymarket-Daten. Dies ist keine Handelsberatung und spielt keine Rolle bei der Auflösung dieses Marktes. · AktualisiertXiaomi 36%
OpenAI 26%
Alibaba 25%
Amazon 24%

Xiaomi
36%

OpenAI
26%

Alibaba
25%

Amazon
24%

Z.ai
24%

SpaceXAI
24%

MiniMax
24%

Anthropic
23%

23%

Meta
14%

ByteDance
14%

Microsoft
14%

Baidu
14%

DeepSeek
13%

Moonshot
12%

Meituan
12%

Tencent
8%

Nvidia
8%

Mistral
7%
Xiaomi 36%
OpenAI 26%
Alibaba 25%
Amazon 24%

Xiaomi
36%

OpenAI
26%

Alibaba
25%

Amazon
24%

Z.ai
24%

SpaceXAI
24%

MiniMax
24%

Anthropic
23%

23%

Meta
14%

ByteDance
14%

Microsoft
14%

Baidu
14%

DeepSeek
13%

Moonshot
12%

Meituan
12%

Tencent
8%

Nvidia
8%

Mistral
7%
Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Models" will be used to resolve this market.
Note: Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the Agent Arena Leaderboard found at https://arena.ai/leaderboard/agent. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve based on another resolution source.
Markt eröffnet: Sep 17, 2026, 8:02 PM ET
Abwickler
0x69c47De9D...Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Models" will be used to resolve this market.
Note: Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the Agent Arena Leaderboard found at https://arena.ai/leaderboard/agent. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve based on another resolution source.
Abwickler
0x69c47De9D...Recent September 2026 model releases have fragmented the AI agent landscape, with no single leader ahead of the November deadline. Anthropic’s Claude Fable 5.1 and Mythos 5.1 lead long-horizon autonomy and Terminal-Bench scores, while OpenAI’s GPT-6 Astra excels on ARC-AGI, ExploitBench, and computer-use tasks; its new Agents API further boosts practical deployment via managed multi-agent workflows and sandboxes. Open-weight contenders from Z.ai (GLM), DeepSeek, and Alibaba (Qwen) post competitive agentic benchmarks at lower cost, sustaining trader parity. Safety coordination among OpenAI, Anthropic, and Google adds regulatory uncertainty, keeping implied probabilities closely matched across labs.
Experimentelle KI-generierte Zusammenfassung mit Polymarket-Daten. Dies ist keine Handelsberatung und spielt keine Rolle bei der Auflösung dieses Marktes. · Aktualisiert
Vorsicht bei externen Links.
Vorsicht bei externen Links.
Häufig gestellte Fragen