Recent releases have sharpened trader focus on MathArena's evolving leaderboards, which now emphasize research-level ArXiv problems, proof generation in Lean, and harder final-answer contests rather than saturated older sets like MATH-500. OpenAI's GPT-6 Astra, launched September 4, 2026, tops expected performance at 91% across non-deprecated tracks, outpacing Anthropic's Claude-Opus-5 and Fable 5.1 models near 72%. Chinese labs continue rapid iteration with Qwen3.8-Max and Kimi variants posting competitive scores on specific subsets, while formal proof and open-problem tracks remain far from saturated. Key catalysts ahead include potential year-end model drops from OpenAI, Anthropic, Google, and xAI that could lift aggregate capabilities before December 31 resolution. Traders weigh these demonstrated gains against historical timeline slippage and the benchmark's continuous refresh of harder problems.
Experimentelle KI-generierte Zusammenfassung mit Polymarket-Daten. Dies ist keine Handelsberatung und spielt keine Rolle bei der Auflösung dieses Marktes. · Aktualisiert$120,652 Vol.
1575
70%
1600
33%
$120,652 Vol.
1575
70%
1600
33%
Results from the "Score" column under the "Text Arena | Math" Leaderboard tab at https://arena.ai/leaderboard/text/math-no-style-control with style control off will be used to resolve this market.
The resolution source for this market is the Chatbot Arena LLM Leaderboard found at arena.ai/leaderboard/text. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If permanently unavailable, this market will resolve to "No".
Markt eröffnet: Apr 2, 2026, 6:07 PM ET
Abwickler
0x65070BE91...Results from the "Score" column under the "Text Arena | Math" Leaderboard tab at https://arena.ai/leaderboard/text/math-no-style-control with style control off will be used to resolve this market.
The resolution source for this market is the Chatbot Arena LLM Leaderboard found at arena.ai/leaderboard/text. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If permanently unavailable, this market will resolve to "No".
Abwickler
0x65070BE91...Recent releases have sharpened trader focus on MathArena's evolving leaderboards, which now emphasize research-level ArXiv problems, proof generation in Lean, and harder final-answer contests rather than saturated older sets like MATH-500. OpenAI's GPT-6 Astra, launched September 4, 2026, tops expected performance at 91% across non-deprecated tracks, outpacing Anthropic's Claude-Opus-5 and Fable 5.1 models near 72%. Chinese labs continue rapid iteration with Qwen3.8-Max and Kimi variants posting competitive scores on specific subsets, while formal proof and open-problem tracks remain far from saturated. Key catalysts ahead include potential year-end model drops from OpenAI, Anthropic, Google, and xAI that could lift aggregate capabilities before December 31 resolution. Traders weigh these demonstrated gains against historical timeline slippage and the benchmark's continuous refresh of harder problems.
Experimentelle KI-generierte Zusammenfassung mit Polymarket-Daten. Dies ist keine Handelsberatung und spielt keine Rolle bei der Auflösung dieses Marktes. · Aktualisiert



Vorsicht bei externen Links.
Vorsicht bei externen Links.
Häufig gestellte Fragen