Skip to main content
icon for Which company has the best AI model on LiveBench (Mathematics) end of August?

Which company has the best AI model on LiveBench (Mathematics) end of August?

icon for Which company has the best AI model on LiveBench (Mathematics) end of August?

Which company has the best AI model on LiveBench (Mathematics) end of August?

OpenAI 45%

Anthropic 38%

DeepSeek 19%

Xiaomi 19%

Polymarket
NOWE

OpenAI 45%

Anthropic 38%

DeepSeek 19%

Xiaomi 19%

Polymarket
NOWE
icon for OpenAI

OpenAI

$128 Wol.

45%

icon for Anthropic

Anthropic

$106 Wol.

27%

icon for DeepSeek

DeepSeek

$67 Wol.

19%

icon for Xiaomi

Xiaomi

$67 Wol.

19%

icon for ByteDance

ByteDance

$89 Wol.

17%

icon for Moonshot

Moonshot

$67 Wol.

13%

icon for SpaceXAI

SpaceXAI

$67 Wol.

13%

icon for Alibaba

Alibaba

$67 Wol.

12%

icon for Meta

Meta

$67 Wol.

12%

icon for Google

Google

$67 Wol.

10%

icon for Baidu

Baidu

$67 Wol.

8%

icon for Z.ai

Z.ai

$67 Wol.

8%

icon for Thinky

Thinky

$138 Wol.

6%

icon for Nvidia

Nvidia

$138 Wol.

7%

icon for Mistral

Mistral

$75 Wol.

3%

icon for Tencent

Tencent

$75 Wol.

3%

icon for Amazon

Amazon

$81 Wol.

3%

icon for Meituan

Meituan

$83 Wol.

1%

icon for StepFun

StepFun

$85 Wol.

1%

icon for Microsoft

Microsoft

$92 Wol.

1%

icon for MiniMax

MiniMax

$92 Wol.

1%

This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”Recent LiveBench mathematics results reflect a tight race among frontier large language models, with market odds clustering near 45-50% for leading contenders as traders weigh incremental gains from reasoning-focused releases. OpenAI's GPT-5 series and o-series variants currently post the strongest scores on the contamination-resistant math tasks, driven by scaled test-time compute and targeted fine-tuning, while Anthropic's Claude 5 Opus and related reasoning modes trail but remain competitive on proof-style and competition problems. Chinese labs like DeepSeek and Moonshot trail further due to narrower gaps in advanced contest math, though rapid iteration keeps them in play. With resolution at month-end, any new model drops, efficiency updates, or benchmark refreshes in August could shift the narrow margins, underscoring how live evaluation rewards timely capability jumps over static leaderboards.

This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET.

Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market.

Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.

The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
Wolumen
$1,785
Data zakończenia
Aug 31, 2026
Rynek otwarty
Jul 29, 2026, 6:27 PM ET
This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”Recent LiveBench mathematics results reflect a tight race among frontier large language models, with market odds clustering near 45-50% for leading contenders as traders weigh incremental gains from reasoning-focused releases. OpenAI's GPT-5 series and o-series variants currently post the strongest scores on the contamination-resistant math tasks, driven by scaled test-time compute and targeted fine-tuning, while Anthropic's Claude 5 Opus and related reasoning modes trail but remain competitive on proof-style and competition problems. Chinese labs like DeepSeek and Moonshot trail further due to narrower gaps in advanced contest math, though rapid iteration keeps them in play. With resolution at month-end, any new model drops, efficiency updates, or benchmark refreshes in August could shift the narrow margins, underscoring how live evaluation rewards timely capability jumps over static leaderboards.

This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET.

Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market.

Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.

The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
Wolumen
$1,785
Data zakończenia
Aug 31, 2026
Rynek otwarty
Jul 29, 2026, 6:27 PM ET
This market will resolve according to the company which owns the model with the highest Mathematics score on LiveBench.ai when the leaderboard is checked on August 31, 2026, at 12:00 PM ET. Results from the “Mathematics” column of the leaderboard at https://livebench.ai/#/?cats=Mathematics, with the latest available LiveBench release selected and the category set to “Mathematics,” will be used to resolve this market. Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking. The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”

Uważaj na linki zewnętrzne.

Często zadawane pytania

"Which company has the best AI model on LiveBench (Mathematics) end of August?" to rynek prognoz na Polymarket z 21 możliwymi wynikami, gdzie traderzy kupują i sprzedają udziały na podstawie tego, co ich zdaniem się wydarzy. Obecny wiodący wynik to "OpenAI" z 45%, za nim "Anthropic" z 27%. Ceny odzwierciedlają zbiorowe prawdopodobieństwa w czasie rzeczywistym. Na przykład udział wyceniony na 45¢ implikuje, że rynek zbiorowo przypisuje 45% szansy na ten wynik. Te kursy zmieniają się ciągle, gdy traderzy reagują na nowe informacje. Udziały w poprawnym wyniku można wymienić na $1 za sztukę po rozstrzygnięciu rynku.

"Which company has the best AI model on LiveBench (Mathematics) end of August?" to nowo utworzony rynek na Polymarket, uruchomiony Jul 29, 2026. Jako wczesny rynek, to Twoja okazja, aby być jednym z pierwszych traderów, którzy ustalą kursy i określą początkowe sygnały cenowe rynku. Możesz też dodać tę stronę do zakładek, aby śledzić wolumen i aktywność handlową w miarę rozwoju rynku.

Aby handlować na "Which company has the best AI model on LiveBench (Mathematics) end of August?", przeglądaj 21 dostępnych wyników na tej stronie. Każdy wynik wyświetla bieżącą cenę reprezentującą implikowane prawdopodobieństwo rynku. Aby zająć pozycję, wybierz wynik, który uważasz za najbardziej prawdopodobny, wybierz "Tak", aby handlować na jego korzyść, lub "Nie", aby handlować przeciw niemu, wpisz kwotę i kliknij "Handluj". Jeśli wybrany wynik okaże się poprawny, Twoje udziały "Tak" wypłacą $1 za sztukę. Jeśli jest niepoprawny, wypłacą $0. Możesz też sprzedać swoje udziały w dowolnym momencie przed rozstrzygnięciem.

Obecnym faworytem dla "Which company has the best AI model on LiveBench (Mathematics) end of August?" jest "OpenAI" z 45%, co oznacza, że rynek przypisuje 45% szansy na ten wynik. Następny najbliższy wynik to "Anthropic" z 27%. Te kursy aktualizują się w czasie rzeczywistym, gdy traderzy kupują i sprzedają udziały, odzwierciedlając najnowszy zbiorowy pogląd na to, co jest najbardziej prawdopodobne. Sprawdzaj regularnie lub dodaj tę stronę do zakładek, aby śledzić zmiany kursów.

Zasady rozstrzygania "Which company has the best AI model on LiveBench (Mathematics) end of August?" określają dokładnie, co musi się wydarzyć, aby każdy wynik został ogłoszony zwycięzcą — w tym oficjalne źródła danych używane do ustalenia wyniku. Możesz przejrzeć pełne kryteria rozstrzygania w sekcji "Zasady" na tej stronie nad komentarzami. Zalecamy dokładne zapoznanie się z zasadami przed handlem, ponieważ określają one precyzyjne warunki, przypadki graniczne i źródła regulujące rozstrzyganie tego rynku.