Recent frontier releases from Anthropic, OpenAI, and Google DeepMind have driven Chatbot Arena Elo scores into the 1500–1531 range as of early September 2026, with Claude Fable 5 and Opus 5 variants frequently leading overall text leaderboards while specialized variants top coding and agent arenas. Post-training refinements, larger context windows, and agentic capabilities continue to lift human-preference ratings, though gains are incremental amid tight clustering among the top five to six labs. Chinese models like Qwen3.8-Max and Kimi K3 add competitive pressure on cost and niche tasks. With multiple expected model updates before year-end, including potential next iterations from OpenAI and Anthropic, trader sentiment hinges on whether incremental improvements or a breakthrough release can push any single model past the market's specified threshold by December 31.
Tóm tắt AI thử nghiệm tham chiếu dữ liệu Polymarket. Đây không phải tư vấn giao dịch và không ảnh hưởng đến cách thị trường này được giải quyết. · Cập nhật$183,515 KL.
↑ 1520
95%
↑ 1530
40%
↑ 1540
20%
↑ 1550
12%
↑ 1600
5%
↑ 1650
5%
↑ 1700
3%
$183,515 KL.
↑ 1520
95%
↑ 1530
40%
↑ 1540
20%
↑ 1550
12%
↑ 1600
5%
↑ 1650
5%
↑ 1700
3%
Results from the 'Score' section on the 'Text Arena' Leaderboard tab (https://lmarena.ai/leaderboard/text), with the style control unchecked, will be used to resolve this market.
The resolution source is the Chatbot Arena LLM Leaderboard (https://lmarena.ai/). If this source is temporarily unavailable, the market remains open until it is accessible again; if permanently unavailable, this market will resolve to "No".
Thị trường mở: Jul 23, 2026, 5:38 PM ET
Người giải quyết
0x65070BE91...Results from the 'Score' section on the 'Text Arena' Leaderboard tab (https://lmarena.ai/leaderboard/text), with the style control unchecked, will be used to resolve this market.
The resolution source is the Chatbot Arena LLM Leaderboard (https://lmarena.ai/). If this source is temporarily unavailable, the market remains open until it is accessible again; if permanently unavailable, this market will resolve to "No".
Người giải quyết
0x65070BE91...Recent frontier releases from Anthropic, OpenAI, and Google DeepMind have driven Chatbot Arena Elo scores into the 1500–1531 range as of early September 2026, with Claude Fable 5 and Opus 5 variants frequently leading overall text leaderboards while specialized variants top coding and agent arenas. Post-training refinements, larger context windows, and agentic capabilities continue to lift human-preference ratings, though gains are incremental amid tight clustering among the top five to six labs. Chinese models like Qwen3.8-Max and Kimi K3 add competitive pressure on cost and niche tasks. With multiple expected model updates before year-end, including potential next iterations from OpenAI and Anthropic, trader sentiment hinges on whether incremental improvements or a breakthrough release can push any single model past the market's specified threshold by December 31.
Tóm tắt AI thử nghiệm tham chiếu dữ liệu Polymarket. Đây không phải tư vấn giao dịch và không ảnh hưởng đến cách thị trường này được giải quyết. · Cập nhật


Cẩn thận với liên kết bên ngoài.
Cẩn thận với liên kết bên ngoài.
Câu hỏi thường gặp