Recent September 2026 model releases and agentic benchmark results have fragmented trader sentiment across frontier labs, with no single clear second-place contender behind the presumed leader. OpenAI’s GPT-6 Astra and Anthropic’s Claude Fable 5.1/Mythos variants posted top verified scores on Terminal-Bench and OSWorld-style evaluations for multi-step tool use and coding agents, while Google’s Gemini 3.8 Flash emphasized speed and cost efficiency in enterprise workflows. Chinese labs including Moonshot, DeepSeek, and Alibaba’s Qwen series remain competitive on open-weight performance and lower inference costs, sustaining uncertainty over which platform will demonstrate superior long-horizon agent capabilities by late November. Upcoming developer conferences and new benchmark drops could shift the closely matched odds.
Resumen experimental generado por IA con datos de Polymarket. Esto no es asesoramiento de trading y no influye en cómo se resuelve este mercado. · ActualizadoBaidu 33.9%
Meta 27%
Microsoft 27%
Mistral 26%

Baidu
34%

Meta
27%

Microsoft
27%

Mistral
26%

DeepSeek
25%

Xiaomi
25%

MiniMax
25%

25%

Amazon
18%

SpaceXAI
14%

ByteDance
14%

Z.ai
14%

Alibaba
14%

Moonshot
14%

Tencent
12%

Nvidia
9%

Meituan
4%

Anthropic
33%

OpenAI
35%
Baidu 33.9%
Meta 27%
Microsoft 27%
Mistral 26%

Baidu
34%

Meta
27%

Microsoft
27%

Mistral
26%

DeepSeek
25%

Xiaomi
25%

MiniMax
25%

25%

Amazon
18%

SpaceXAI
14%

ByteDance
14%

Z.ai
14%

Alibaba
14%

Moonshot
14%

Tencent
12%

Nvidia
9%

Meituan
4%

Anthropic
33%

OpenAI
35%
Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Labs" will be used to resolve this market.
Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
AI companies will be ordered primarily by their Lab Rank at the market’s check time. If the results based on the lab ranking are ambiguous or unavailable, the relevant AI companies will be ordered according to their highest-ranking AI model in the leaderboard’s “Models” view. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of AI lab/company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies second place under this ranking.
The resolution source for this market is the arena.ai Agent Arena Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Mercado abierto: Sep 17, 2026, 8:03 PM ET
Resolver
0x69c47De9D...Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Labs" will be used to resolve this market.
Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
AI companies will be ordered primarily by their Lab Rank at the market’s check time. If the results based on the lab ranking are ambiguous or unavailable, the relevant AI companies will be ordered according to their highest-ranking AI model in the leaderboard’s “Models” view. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of AI lab/company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies second place under this ranking.
The resolution source for this market is the arena.ai Agent Arena Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Resolver
0x69c47De9D...Recent September 2026 model releases and agentic benchmark results have fragmented trader sentiment across frontier labs, with no single clear second-place contender behind the presumed leader. OpenAI’s GPT-6 Astra and Anthropic’s Claude Fable 5.1/Mythos variants posted top verified scores on Terminal-Bench and OSWorld-style evaluations for multi-step tool use and coding agents, while Google’s Gemini 3.8 Flash emphasized speed and cost efficiency in enterprise workflows. Chinese labs including Moonshot, DeepSeek, and Alibaba’s Qwen series remain competitive on open-weight performance and lower inference costs, sustaining uncertainty over which platform will demonstrate superior long-horizon agent capabilities by late November. Upcoming developer conferences and new benchmark drops could shift the closely matched odds.
Resumen experimental generado por IA con datos de Polymarket. Esto no es asesoramiento de trading y no influye en cómo se resuelve este mercado. · Actualizado
Cuidado con los enlaces externos.
Cuidado con los enlaces externos.
Preguntas frecuentes