Recent releases from frontier labs have driven rapid gains on MathArena, the contamination-resistant benchmark tracking LLM performance on fresh math olympiad and research problems. OpenAI’s GPT-6 Astra, launched in early September 2026, currently leads with expected performance near 90.7 percent, ahead of Anthropic’s Claude-Fable-5.1 and Opus-5 variants. Labs continue iterating on reasoning techniques and scale, with ongoing updates to the leaderboard reflecting new evaluations on AIME, Putnam-style contests, and arXiv problems. Traders are watching for additional model drops or capability jumps before year-end that could push the top score past key thresholds amid intense competition among OpenAI, Anthropic, Google, and xAI.
Экспериментальная сводка, созданная ИИ на основе данных Polymarket. Это не является торговой рекомендацией и не влияет на то, как разрешается этот рынок. · ОбновленоWill any AI model reach ___ Math Arena Score by December 31?
$126,094 Объем
1575
48%
1600
9%
$126,094 Объем
1575
48%
1600
9%
Results from the "Score" column under the "Text Arena | Math" Leaderboard tab at https://arena.ai/leaderboard/text/math-no-style-control with style control off will be used to resolve this market.
The resolution source for this market is the Chatbot Arena LLM Leaderboard found at arena.ai/leaderboard/text. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If permanently unavailable, this market will resolve to "No".
Открытие рынка: Apr 2, 2026, 6:07 PM ET
Кто определяет исход
0x65070BE91...Results from the "Score" column under the "Text Arena | Math" Leaderboard tab at https://arena.ai/leaderboard/text/math-no-style-control with style control off will be used to resolve this market.
The resolution source for this market is the Chatbot Arena LLM Leaderboard found at arena.ai/leaderboard/text. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If permanently unavailable, this market will resolve to "No".
Кто определяет исход
0x65070BE91...Recent releases from frontier labs have driven rapid gains on MathArena, the contamination-resistant benchmark tracking LLM performance on fresh math olympiad and research problems. OpenAI’s GPT-6 Astra, launched in early September 2026, currently leads with expected performance near 90.7 percent, ahead of Anthropic’s Claude-Fable-5.1 and Opus-5 variants. Labs continue iterating on reasoning techniques and scale, with ongoing updates to the leaderboard reflecting new evaluations on AIME, Putnam-style contests, and arXiv problems. Traders are watching for additional model drops or capability jumps before year-end that could push the top score past key thresholds amid intense competition among OpenAI, Anthropic, Google, and xAI.
Экспериментальная сводка, созданная ИИ на основе данных Polymarket. Это не является торговой рекомендацией и не влияет на то, как разрешается этот рынок. · Обновлено



Не доверяй внешним ссылкам.
Не доверяй внешним ссылкам.
Часто задаваемые вопросы