Recent benchmark leadership in agentic tasks such as Terminal-Bench, OSWorld-Verified, and BrowseComp remains fragmented, with Shanghai AI Lab’s Atria Dawn Preview, OpenAI’s GPT-5.6/GPT-6 series, Moonshot’s Kimi K3, Anthropic’s Claude Opus 5, and DeepSeek models trading top spots across coding, browsing, and multi-step workflows. This fluidity, combined with strong revenue traction at Cognition (Devin) and Sierra alongside rapid open-weight progress from Z.ai, Alibaba Qwen, and Moonshot, sustains broad trader uncertainty over which lab will rank third by late November. Differentiators include demonstrated long-horizon reliability, enterprise deployment scale, and post-training optimizations for tool use rather than raw model size. Upcoming model releases and benchmark updates through October could quickly shift the crowded field.
Polymarket ডেটা রেফারেন্স করে পরীক্ষামূলক AI-জেনারেটেড সারাংশ। এটি ট্রেডিং পরামর্শ নয় এবং এই মার্কেট কীভাবে রেজলভ হয় তাতে কোনো ভূমিকা রাখে না। · আপডেটেডMoonshot 28%
Z.ai 22%
Google 22%
ByteDance 20%

Moonshot
28%

Z.ai
22%

22%

ByteDance
20%

Tencent
20%

SpaceXAI
18%

DeepSeek
18%

Anthropic
14%

OpenAI
14%

Baidu
12%

Amazon
12%

Alibaba
12%

Nvidia
12%

Mistral
12%

Meituan
11%

Microsoft
11%

Meta
9%

Xiaomi
8%

MiniMax
8%
Moonshot 28%
Z.ai 22%
Google 22%
ByteDance 20%

Moonshot
28%

Z.ai
22%

22%

ByteDance
20%

Tencent
20%

SpaceXAI
18%

DeepSeek
18%

Anthropic
14%

OpenAI
14%

Baidu
12%

Amazon
12%

Alibaba
12%

Nvidia
12%

Mistral
12%

Meituan
11%

Microsoft
11%

Meta
9%

Xiaomi
8%

MiniMax
8%
Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Labs" will be used to resolve this market.
Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
AI companies will be ordered primarily by their Lab Rank at the market’s check time. If the results based on the lab ranking are ambiguous or unavailable, the relevant AI companies will be ordered according to their highest-ranking AI model in the leaderboard’s “Models” view. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of AI lab/company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies third place under this ranking.
The resolution source for this market is the arena.ai Agent Arena Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
মার্কেট ওপেন হয়েছে: Sep 17, 2026, 8:04 PM ET
রেজলভার
0x69c47De9D...Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Labs" will be used to resolve this market.
Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
AI companies will be ordered primarily by their Lab Rank at the market’s check time. If the results based on the lab ranking are ambiguous or unavailable, the relevant AI companies will be ordered according to their highest-ranking AI model in the leaderboard’s “Models” view. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of AI lab/company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies third place under this ranking.
The resolution source for this market is the arena.ai Agent Arena Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
রেজলভার
0x69c47De9D...Recent benchmark leadership in agentic tasks such as Terminal-Bench, OSWorld-Verified, and BrowseComp remains fragmented, with Shanghai AI Lab’s Atria Dawn Preview, OpenAI’s GPT-5.6/GPT-6 series, Moonshot’s Kimi K3, Anthropic’s Claude Opus 5, and DeepSeek models trading top spots across coding, browsing, and multi-step workflows. This fluidity, combined with strong revenue traction at Cognition (Devin) and Sierra alongside rapid open-weight progress from Z.ai, Alibaba Qwen, and Moonshot, sustains broad trader uncertainty over which lab will rank third by late November. Differentiators include demonstrated long-horizon reliability, enterprise deployment scale, and post-training optimizations for tool use rather than raw model size. Upcoming model releases and benchmark updates through October could quickly shift the crowded field.
Polymarket ডেটা রেফারেন্স করে পরীক্ষামূলক AI-জেনারেটেড সারাংশ। এটি ট্রেডিং পরামর্শ নয় এবং এই মার্কেট কীভাবে রেজলভ হয় তাতে কোনো ভূমিকা রাখে না। · আপডেটেড
বাহ্যিক লিংক থেকে সাবধান।
বাহ্যিক লিংক থেকে সাবধান।
সচরাচর জিজ্ঞাসা