Skip to main content
icon for Which company has the best AI model on LiveBench (Mathematics) end of August?

Which company has the best AI model on LiveBench (Mathematics) end of August?

icon for Which company has the best AI model on LiveBench (Mathematics) end of August?

Which company has the best AI model on LiveBench (Mathematics) end of August?

OpenAI 45%

Anthropic 38%

Alibaba 28%

DeepSeek 22%

Polymarket
NOUVEAU

OpenAI 45%

Anthropic 38%

Alibaba 28%

DeepSeek 22%

Polymarket
NOUVEAU
icon for OpenAI

OpenAI

$128 Vol.

43%

icon for Anthropic

Anthropic

$106 Vol.

34%

icon for Alibaba

Alibaba

$67 Vol.

28%

icon for DeepSeek

DeepSeek

$67 Vol.

22%

icon for Xiaomi

Xiaomi

$67 Vol.

22%

icon for SpaceXAI

SpaceXAI

$67 Vol.

22%

icon for Google

Google

$67 Vol.

22%

icon for Meta

Meta

$67 Vol.

21%

icon for ByteDance

ByteDance

$89 Vol.

21%

icon for Moonshot

Moonshot

$67 Vol.

20%

icon for Baidu

Baidu

$67 Vol.

9%

icon for Z.ai

Z.ai

$67 Vol.

9%

icon for Thinky

Thinky

$138 Vol.

6%

icon for Nvidia

Nvidia

$138 Vol.

7%

icon for Mistral

Mistral

$75 Vol.

3%

icon for Tencent

Tencent

$75 Vol.

3%

icon for Amazon

Amazon

$81 Vol.

3%

icon for Meituan

Meituan

$83 Vol.

1%

icon for StepFun

StepFun

$85 Vol.

1%

icon for Microsoft

Microsoft

$92 Vol.

1%

icon for MiniMax

MiniMax

$92 Vol.

-

This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”Recent LiveBench mathematics results reflect a tight race among frontier large language models, with market odds clustering near 45-50% for leading contenders as traders weigh incremental gains from reasoning-focused releases. OpenAI's GPT-5 series and o-series variants currently post the strongest scores on the contamination-resistant math tasks, driven by scaled test-time compute and targeted fine-tuning, while Anthropic's Claude 5 Opus and related reasoning modes trail but remain competitive on proof-style and competition problems. Chinese labs like DeepSeek and Moonshot trail further due to narrower gaps in advanced contest math, though rapid iteration keeps them in play. With resolution at month-end, any new model drops, efficiency updates, or benchmark refreshes in August could shift the narrow margins, underscoring how live evaluation rewards timely capability jumps over static leaderboards.

This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET.

Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market.

Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.

The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
Volume
$1,785
Date de fin
31 août 2026
Marché ouvert
Jul 29, 2026, 6:27 PM ET
This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”Recent LiveBench mathematics results reflect a tight race among frontier large language models, with market odds clustering near 45-50% for leading contenders as traders weigh incremental gains from reasoning-focused releases. OpenAI's GPT-5 series and o-series variants currently post the strongest scores on the contamination-resistant math tasks, driven by scaled test-time compute and targeted fine-tuning, while Anthropic's Claude 5 Opus and related reasoning modes trail but remain competitive on proof-style and competition problems. Chinese labs like DeepSeek and Moonshot trail further due to narrower gaps in advanced contest math, though rapid iteration keeps them in play. With resolution at month-end, any new model drops, efficiency updates, or benchmark refreshes in August could shift the narrow margins, underscoring how live evaluation rewards timely capability jumps over static leaderboards.

This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET.

Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market.

Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.

The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
Volume
$1,785
Date de fin
31 août 2026
Marché ouvert
Jul 29, 2026, 6:27 PM ET
This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”

Méfiez-vous des liens externes.

Questions fréquentes

« Which company has the best AI model on LiveBench (Mathematics) end of August? » est un marché de prédiction sur Polymarket avec 21 résultats possibles où les traders achètent et vendent des parts selon ce qu'ils pensent qu'il se passera. Le résultat en tête actuel est « OpenAI » à 43%, suivi de « Anthropic » à 34%. Les prix reflètent des probabilités en temps réel de la communauté. Par exemple, une part cotée à 43¢ implique que le marché attribue collectivement une probabilité de 43% à ce résultat. Ces cotes changent en permanence. Les parts du résultat correct sont échangeables contre $1 chacune lors de la résolution du marché.

« Which company has the best AI model on LiveBench (Mathematics) end of August? » est un marché nouvellement créé sur Polymarket, lancé le Jul 29, 2026. En tant que marché récent, c'est votre opportunité d'être parmi les premiers traders à définir les cotes et établir les premiers signaux de prix du marché. Vous pouvez également ajouter cette page à vos favoris pour suivre le volume et l'activité de trading au fil du temps.

Pour trader sur « Which company has the best AI model on LiveBench (Mathematics) end of August? », parcourez les 21 résultats disponibles sur cette page. Chaque résultat affiche un prix actuel représentant la probabilité implicite du marché. Pour prendre position, sélectionnez le résultat que vous estimez le plus probable, choisissez « Oui » pour trader en sa faveur ou « Non » pour trader contre, entrez votre montant et cliquez sur « Trader ». Si votre résultat choisi est correct lors de la résolution, vos parts « Oui » rapportent $1 chacune. S'il est incorrect, elles rapportent $0. Vous pouvez également vendre vos parts avant la résolution.

Le favori actuel pour « Which company has the best AI model on LiveBench (Mathematics) end of August? » est « OpenAI » à 43%, ce qui signifie que le marché attribue une probabilité de 43% à ce résultat. Le résultat le plus proche ensuite est « Anthropic » à 34%. Ces cotes sont mises à jour en temps réel à mesure que les traders achètent et vendent des parts. Revenez fréquemment ou ajoutez cette page à vos favoris.

Les règles de résolution de « Which company has the best AI model on LiveBench (Mathematics) end of August? » définissent exactement ce qui doit se produire pour que chaque résultat soit déclaré gagnant, y compris les sources de données officielles utilisées pour déterminer le résultat. Vous pouvez consulter les critères de résolution complets dans la section « Règles » sur cette page au-dessus des commentaires. Nous recommandons de lire attentivement les règles avant de trader, car elles précisent les conditions exactes, les cas particuliers et les sources.