Skip to main content
icon for Which company has the best AI model on LiveBench (Mathematics) end of August?

Which company has the best AI model on LiveBench (Mathematics) end of August?

icon for Which company has the best AI model on LiveBench (Mathematics) end of August?

Which company has the best AI model on LiveBench (Mathematics) end of August?

OpenAI 45%

Anthropic 38%

Alibaba 27%

DeepSeek 22%

Polymarket
NUEVO

OpenAI 45%

Anthropic 38%

Alibaba 27%

DeepSeek 22%

Polymarket
NUEVO
icon for OpenAI

OpenAI

$128 Vol.

43%

icon for Anthropic

Anthropic

$106 Vol.

34%

icon for Alibaba

Alibaba

$67 Vol.

27%

icon for DeepSeek

DeepSeek

$67 Vol.

22%

icon for Xiaomi

Xiaomi

$67 Vol.

22%

icon for SpaceXAI

SpaceXAI

$67 Vol.

22%

icon for Google

Google

$67 Vol.

22%

icon for Moonshot

Moonshot

$67 Vol.

21%

icon for Meta

Meta

$67 Vol.

21%

icon for ByteDance

ByteDance

$89 Vol.

19%

icon for Z.ai

Z.ai

$67 Vol.

18%

icon for Baidu

Baidu

$67 Vol.

9%

icon for Thinky

Thinky

$138 Vol.

6%

icon for Nvidia

Nvidia

$138 Vol.

7%

icon for Mistral

Mistral

$75 Vol.

3%

icon for Tencent

Tencent

$75 Vol.

3%

icon for Amazon

Amazon

$81 Vol.

3%

icon for Meituan

Meituan

$83 Vol.

1%

icon for StepFun

StepFun

$85 Vol.

1%

icon for Microsoft

Microsoft

$92 Vol.

1%

icon for MiniMax

MiniMax

$92 Vol.

-

This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”Recent LiveBench mathematics results reflect a tight race among frontier large language models, with market odds clustering near 45-50% for leading contenders as traders weigh incremental gains from reasoning-focused releases. OpenAI's GPT-5 series and o-series variants currently post the strongest scores on the contamination-resistant math tasks, driven by scaled test-time compute and targeted fine-tuning, while Anthropic's Claude 5 Opus and related reasoning modes trail but remain competitive on proof-style and competition problems. Chinese labs like DeepSeek and Moonshot trail further due to narrower gaps in advanced contest math, though rapid iteration keeps them in play. With resolution at month-end, any new model drops, efficiency updates, or benchmark refreshes in August could shift the narrow margins, underscoring how live evaluation rewards timely capability jumps over static leaderboards.

This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET.

Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market.

Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.

The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
Volumen
$1,785
Fecha de finalización
31 ago 2026
Mercado abierto
Jul 29, 2026, 6:27 PM ET
This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”Recent LiveBench mathematics results reflect a tight race among frontier large language models, with market odds clustering near 45-50% for leading contenders as traders weigh incremental gains from reasoning-focused releases. OpenAI's GPT-5 series and o-series variants currently post the strongest scores on the contamination-resistant math tasks, driven by scaled test-time compute and targeted fine-tuning, while Anthropic's Claude 5 Opus and related reasoning modes trail but remain competitive on proof-style and competition problems. Chinese labs like DeepSeek and Moonshot trail further due to narrower gaps in advanced contest math, though rapid iteration keeps them in play. With resolution at month-end, any new model drops, efficiency updates, or benchmark refreshes in August could shift the narrow margins, underscoring how live evaluation rewards timely capability jumps over static leaderboards.

This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET.

Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market.

Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.

The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
Volumen
$1,785
Fecha de finalización
31 ago 2026
Mercado abierto
Jul 29, 2026, 6:27 PM ET
This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”

Cuidado con los enlaces externos.

Preguntas frecuentes

"Which company has the best AI model on LiveBench (Mathematics) end of August?" es un mercado de predicción en Polymarket con 21 resultados posibles donde los operadores compran y venden acciones según lo que creen que sucederá. El resultado líder actual es "OpenAI" con 43%, seguido de "Anthropic" con 34%. Los precios reflejan probabilidades en tiempo real de la comunidad. Por ejemplo, una acción cotizada a 43¢ implica que el mercado colectivamente asigna una probabilidad de 43% a ese resultado. Estas probabilidades cambian continuamente a medida que los operadores reaccionan a nuevos desarrollos. Las acciones del resultado correcto son canjeables por $1 cada una tras la resolución del mercado.

"Which company has the best AI model on LiveBench (Mathematics) end of August?" es un mercado recién creado en Polymarket, lanzado el Jul 29, 2026. Como mercado nuevo, esta es tu oportunidad de ser uno de los primeros operadores en establecer las probabilidades y las señales de precio iniciales del mercado. También puedes guardar esta página en marcadores para seguir el volumen y la actividad de trading a medida que el mercado gana tracción.

Para operar en "Which company has the best AI model on LiveBench (Mathematics) end of August?", explora los 21 resultados disponibles en esta página. Cada resultado muestra un precio actual que representa la probabilidad implícita del mercado. Para tomar una posición, selecciona el resultado que consideres más probable, elige "Sí" para operar a favor o "No" para operar en contra, introduce tu cantidad y haz clic en "Operar". Si tu resultado elegido es correcto cuando el mercado se resuelve, tus acciones de "Sí" pagan $1 cada una. Si es incorrecto, pagan $0. También puedes vender tus acciones en cualquier momento antes de la resolución.

El favorito actual para "Which company has the best AI model on LiveBench (Mathematics) end of August?" es "OpenAI" con 43%, lo que significa que el mercado asigna una probabilidad de 43% a ese resultado. El siguiente resultado más cercano es "Anthropic" con 34%. Estas probabilidades se actualizan en tiempo real a medida que los operadores compran y venden acciones. Vuelve con frecuencia o guarda esta página en marcadores.

Las reglas de resolución para "Which company has the best AI model on LiveBench (Mathematics) end of August?" definen exactamente qué debe ocurrir para que cada resultado sea declarado ganador, incluyendo las fuentes de datos oficiales utilizadas para determinar el resultado. Puedes revisar los criterios de resolución completos en la sección "Reglas" en esta página sobre los comentarios. Recomendamos leer las reglas cuidadosamente antes de operar, ya que especifican las condiciones exactas, casos especiales y fuentes.