Anthropic’s Claude Opus 5 series leads the LM Arena Coding/WebDev leaderboard with claude-opus-5-max at 1688 Elo as of late August 2026, reflecting strong user-voted performance on multi-step agentic tasks. Moonshot’s Kimi K3-max and Alibaba’s Qwen3.8-max sit close behind at 1674 and 1669, narrowing the gap through rapid iteration and competitive pricing. OpenAI’s GPT-5.6 variants and xAI’s Grok-4.6 models also post high marks on related benchmarks like SWE-bench Verified and Terminal-Bench. With four months remaining, ongoing model updates, scaling, and fine-tuning could push top scores higher before year-end, though leaderboard volatility and new entrants create uncertainty around any specific threshold.
Resumo experimental gerado por IA com dados do Polymarket. Isto não é aconselhamento de trading e não tem qualquer papel na resolução deste mercado. · AtualizadoWill any AI model reach ___ Coding Arena Score by December 31?
$192,345 Vol.
1560
33%
1580
21%
1600
13%
$192,345 Vol.
1560
33%
1580
21%
1600
13%
Results from the "Score" column under the "Text Arena | Coding" Leaderboard tab at https://arena.ai/leaderboard/text/coding-no-style-control with style control off will be used to resolve this market.
The resolution source for this market is the Chatbot Arena LLM Leaderboard found at arena.ai/leaderboard/text. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If permanently unavailable, this market will resolve to "No".
Mercado Aberto: Apr 2, 2026, 6:09 PM ET
Resolver
0x65070BE91...Results from the "Score" column under the "Text Arena | Coding" Leaderboard tab at https://arena.ai/leaderboard/text/coding-no-style-control with style control off will be used to resolve this market.
The resolution source for this market is the Chatbot Arena LLM Leaderboard found at arena.ai/leaderboard/text. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If permanently unavailable, this market will resolve to "No".
Resolver
0x65070BE91...Anthropic’s Claude Opus 5 series leads the LM Arena Coding/WebDev leaderboard with claude-opus-5-max at 1688 Elo as of late August 2026, reflecting strong user-voted performance on multi-step agentic tasks. Moonshot’s Kimi K3-max and Alibaba’s Qwen3.8-max sit close behind at 1674 and 1669, narrowing the gap through rapid iteration and competitive pricing. OpenAI’s GPT-5.6 variants and xAI’s Grok-4.6 models also post high marks on related benchmarks like SWE-bench Verified and Terminal-Bench. With four months remaining, ongoing model updates, scaling, and fine-tuning could push top scores higher before year-end, though leaderboard volatility and new entrants create uncertainty around any specific threshold.
Resumo experimental gerado por IA com dados do Polymarket. Isto não é aconselhamento de trading e não tem qualquer papel na resolução deste mercado. · Atualizado



Cuidado com os links externos.
Cuidado com os links externos.
Frequently Asked Questions