Skip to main content
icon for Which company has the best AI model on LiveBench (Mathematics) end of August?

Which company has the best AI model on LiveBench (Mathematics) end of August?

icon for Which company has the best AI model on LiveBench (Mathematics) end of August?

Which company has the best AI model on LiveBench (Mathematics) end of August?

OpenAI 45%

Anthropic 38%

Alibaba 24%

DeepSeek 22%

Polymarket
最新

OpenAI 45%

Anthropic 38%

Alibaba 24%

DeepSeek 22%

Polymarket
最新
icon for OpenAI

OpenAI

$128 交易量

42%

icon for Anthropic

Anthropic

$106 交易量

32%

icon for Alibaba

Alibaba

$67 交易量

24%

icon for DeepSeek

DeepSeek

$67 交易量

22%

icon for Xiaomi

Xiaomi

$67 交易量

22%

icon for SpaceXAI

SpaceXAI

$67 交易量

22%

icon for Google

Google

$67 交易量

22%

icon for Meta

Meta

$67 交易量

21%

icon for Moonshot

Moonshot

$67 交易量

21%

icon for ByteDance

ByteDance

$89 交易量

18%

icon for Baidu

Baidu

$67 交易量

9%

icon for Z.ai

Z.ai

$67 交易量

9%

icon for Thinky

Thinky

$138 交易量

6%

icon for Nvidia

Nvidia

$138 交易量

7%

icon for Mistral

Mistral

$75 交易量

3%

icon for Tencent

Tencent

$75 交易量

3%

icon for Amazon

Amazon

$81 交易量

3%

icon for Meituan

Meituan

$83 交易量

1%

icon for StepFun

StepFun

$85 交易量

1%

icon for Microsoft

Microsoft

$92 交易量

-

icon for MiniMax

MiniMax

$92 交易量

-

This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”Recent LiveBench mathematics results reflect a tight race among frontier large language models, with market odds clustering near 45-50% for leading contenders as traders weigh incremental gains from reasoning-focused releases. OpenAI's GPT-5 series and o-series variants currently post the strongest scores on the contamination-resistant math tasks, driven by scaled test-time compute and targeted fine-tuning, while Anthropic's Claude 5 Opus and related reasoning modes trail but remain competitive on proof-style and competition problems. Chinese labs like DeepSeek and Moonshot trail further due to narrower gaps in advanced contest math, though rapid iteration keeps them in play. With resolution at month-end, any new model drops, efficiency updates, or benchmark refreshes in August could shift the narrow margins, underscoring how live evaluation rewards timely capability jumps over static leaderboards.

This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET.

Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market.

Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.

The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
交易量
$1,785
结束日期
2026-08-31
市场开放时间
Jul 29, 2026, 6:27 PM ET
This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”Recent LiveBench mathematics results reflect a tight race among frontier large language models, with market odds clustering near 45-50% for leading contenders as traders weigh incremental gains from reasoning-focused releases. OpenAI's GPT-5 series and o-series variants currently post the strongest scores on the contamination-resistant math tasks, driven by scaled test-time compute and targeted fine-tuning, while Anthropic's Claude 5 Opus and related reasoning modes trail but remain competitive on proof-style and competition problems. Chinese labs like DeepSeek and Moonshot trail further due to narrower gaps in advanced contest math, though rapid iteration keeps them in play. With resolution at month-end, any new model drops, efficiency updates, or benchmark refreshes in August could shift the narrow margins, underscoring how live evaluation rewards timely capability jumps over static leaderboards.

This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET.

Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market.

Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.

The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
交易量
$1,785
结束日期
2026-08-31
市场开放时间
Jul 29, 2026, 6:27 PM ET
This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”

警惕外部链接哦。

常见问题

"Which company has the best AI model on LiveBench (Mathematics) end of August?"是 Polymarket 上一个拥有 21 个可能结果的预测市场,交易者根据自己的判断买卖份额。当前领先结果为"OpenAI",概率为 42%,其次是"Anthropic",概率为 32%。价格反映社区的实时概率。例如,价格为 42¢ 的份额意味着市场集体认为该结果的概率为 42%。这些赔率会随着交易者的反应而不断变化。正确结果的份额在市场结算时可兑换为每份 $1。

"Which company has the best AI model on LiveBench (Mathematics) end of August?"是 Polymarket 上新创建的市场,于Jul 29, 2026上线。作为一个新市场,这是你率先设定赔率并建立初始价格信号的机会。你也可以将本页加入书签,以便跟踪交易量和活动。

要在"Which company has the best AI model on LiveBench (Mathematics) end of August?"上交易,浏览本页上列出的 21 个可用结果。每个结果显示一个代表市场隐含概率的当前价格。要建仓,选择你认为最可能的结果,选择"是"支持或"否"反对,输入金额并点击"交易"。如果你选择的结果在市场结算时正确,你的"是"份额每份支付 $1。如果不正确,支付 $0。你也可以在结算前随时卖出份额。

"Which company has the best AI model on LiveBench (Mathematics) end of August?"的当前领先者是"OpenAI",概率为 42%,意味着市场对该结果的概率评估为 42%。紧随其后的结果是"Anthropic",概率为 32%。这些赔率随着交易者买卖份额而实时更新。请经常回来查看或将本页加入书签。

"Which company has the best AI model on LiveBench (Mathematics) end of August?"的结算规则明确定义了每个结果被宣布为获胜者所需满足的条件——包括用于确定结果的官方数据来源。你可以在本页评论上方的"规则"部分查看完整的结算标准。我们建议在交易前仔细阅读规则,因为它们规定了精确的条件、特殊情况和数据来源。