Recent September 2026 model releases and agentic benchmark results have fragmented trader sentiment across frontier labs, with no single clear second-place contender behind the presumed leader. OpenAI’s GPT-6 Astra and Anthropic’s Claude Fable 5.1/Mythos variants posted top verified scores on Terminal-Bench and OSWorld-style evaluations for multi-step tool use and coding agents, while Google’s Gemini 3.8 Flash emphasized speed and cost efficiency in enterprise workflows. Chinese labs including Moonshot, DeepSeek, and Alibaba’s Qwen series remain competitive on open-weight performance and lower inference costs, sustaining uncertainty over which platform will demonstrate superior long-horizon agent capabilities by late November. Upcoming developer conferences and new benchmark drops could shift the closely matched odds.
Resumo experimental gerado por IA com dados do Polymarket. Isto não é aconselhamento de trading e não tem qualquer papel na resolução deste mercado. · AtualizadoSegundo melhor laboratório de agentes de IA no final de novembro?
Baidu 29.7%
Meta 27%
Microsoft 27%
ByteDance 26%

Baidu
30%

Meta
27%

Microsoft
27%

ByteDance
26%

Alibaba
26%

Xiaomi
26%

Mistral
26%

25%

Moonshot
25%

Tencent
24%

Z.ai
14%

SpaceXAI
14%

DeepSeek
14%

MiniMax
14%

Amazon
14%

Nvidia
9%

Meituan
4%

Anthropic
33%

OpenAI
31%
Baidu 29.7%
Meta 27%
Microsoft 27%
ByteDance 26%

Baidu
30%

Meta
27%

Microsoft
27%

ByteDance
26%

Alibaba
26%

Xiaomi
26%

Mistral
26%

25%

Moonshot
25%

Tencent
24%

Z.ai
14%

SpaceXAI
14%

DeepSeek
14%

MiniMax
14%

Amazon
14%

Nvidia
9%

Meituan
4%

Anthropic
33%

OpenAI
31%
Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Labs" will be used to resolve this market.
Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
AI companies will be ordered primarily by their Lab Rank at the market’s check time. If the results based on the lab ranking are ambiguous or unavailable, the relevant AI companies will be ordered according to their highest-ranking AI model in the leaderboard’s “Models” view. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of AI lab/company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies second place under this ranking.
The resolution source for this market is the arena.ai Agent Arena Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Mercado Aberto: Sep 17, 2026, 8:03 PM ET
Resolver
0x69c47De9D...Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Labs" will be used to resolve this market.
Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
AI companies will be ordered primarily by their Lab Rank at the market’s check time. If the results based on the lab ranking are ambiguous or unavailable, the relevant AI companies will be ordered according to their highest-ranking AI model in the leaderboard’s “Models” view. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of AI lab/company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies second place under this ranking.
The resolution source for this market is the arena.ai Agent Arena Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Resolver
0x69c47De9D...Recent September 2026 model releases and agentic benchmark results have fragmented trader sentiment across frontier labs, with no single clear second-place contender behind the presumed leader. OpenAI’s GPT-6 Astra and Anthropic’s Claude Fable 5.1/Mythos variants posted top verified scores on Terminal-Bench and OSWorld-style evaluations for multi-step tool use and coding agents, while Google’s Gemini 3.8 Flash emphasized speed and cost efficiency in enterprise workflows. Chinese labs including Moonshot, DeepSeek, and Alibaba’s Qwen series remain competitive on open-weight performance and lower inference costs, sustaining uncertainty over which platform will demonstrate superior long-horizon agent capabilities by late November. Upcoming developer conferences and new benchmark drops could shift the closely matched odds.
Resumo experimental gerado por IA com dados do Polymarket. Isto não é aconselhamento de trading e não tem qualquer papel na resolução deste mercado. · Atualizado
Cuidado com os links externos.
Cuidado com os links externos.
Frequently Asked Questions