Anthropic leads trader sentiment on the LiveBench Mathematics market at 88.5% implied probability due to its recent Claude variants posting standout results on hard math benchmarks, including strong USAMO 2026 scores and a high-profile formalization of Fermat’s Last Theorem in early September. These outcomes reflect targeted gains in mathematical reasoning within large language models, where LiveBench’s contamination-resistant problems test frontier capabilities more rigorously than older sets. OpenAI trails sharply at 9.5% despite topping overall LiveBench averages with models like o3-mini, as its edge appears narrower in the dedicated Mathematics column. Other labs remain at negligible odds because none have demonstrated comparable recent math-specific lifts. With resolution just weeks away on September 30, any new model drop or leaderboard refresh could still alter the ranking before the final snapshot.
基于Polymarket数据的AI实验性摘要。这不是交易建议,也不影响该市场的结算方式。 · 更新于Anthropic 89%
OpenAI 10%
DeepSeek <1%
Meta <1%
$50,337 交易量
$50,337 交易量

Anthropic
89%

OpenAI
10%

DeepSeek
<1%

Meta
<1%

谷歌
<1%

阿里巴巴
<1%

Moonshot
<1%

Baidu
<1%

Z.ai
<1%

Xiaomi
<1%

SpaceXAI
<1%

Amazon
<1%

ByteDance
<1%

Nvidia
<1%

Thinky
<1%

MiniMax
<1%

Mistral
<1%

Meituan
<1%

腾讯
<1%

StepFun
<1%

微软
<1%
Anthropic 89%
OpenAI 10%
DeepSeek <1%
Meta <1%
$50,337 交易量
$50,337 交易量

Anthropic
89%

OpenAI
10%

DeepSeek
<1%

Meta
<1%

谷歌
<1%

阿里巴巴
<1%

Moonshot
<1%

Baidu
<1%

Z.ai
<1%

Xiaomi
<1%

SpaceXAI
<1%

Amazon
<1%

ByteDance
<1%

Nvidia
<1%

Thinky
<1%

MiniMax
<1%

Mistral
<1%

Meituan
<1%

腾讯
<1%

StepFun
<1%

微软
<1%
Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market.
Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
市场开放时间: Jul 29, 2026, 6:14 PM ET
Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market.
Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
Anthropic leads trader sentiment on the LiveBench Mathematics market at 88.5% implied probability due to its recent Claude variants posting standout results on hard math benchmarks, including strong USAMO 2026 scores and a high-profile formalization of Fermat’s Last Theorem in early September. These outcomes reflect targeted gains in mathematical reasoning within large language models, where LiveBench’s contamination-resistant problems test frontier capabilities more rigorously than older sets. OpenAI trails sharply at 9.5% despite topping overall LiveBench averages with models like o3-mini, as its edge appears narrower in the dedicated Mathematics column. Other labs remain at negligible odds because none have demonstrated comparable recent math-specific lifts. With resolution just weeks away on September 30, any new model drop or leaderboard refresh could still alter the ranking before the final snapshot.
基于Polymarket数据的AI实验性摘要。这不是交易建议,也不影响该市场的结算方式。 · 更新于
警惕外部链接哦。
警惕外部链接哦。
常见问题