Skip to main content
icon for AI model scores ≥ 90% on FrontierMath Benchmark before 2027?

AI model scores ≥ 90% on FrontierMath Benchmark before 2027?

icon for AI model scores ≥ 90% on FrontierMath Benchmark before 2027?

AI model scores ≥ 90% on FrontierMath Benchmark before 2027?

Dec 31

Dec 31

90% шанс
Polymarket

$117,315 Обс.

90% шанс
Polymarket

$117,315 Обс.

This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used. **Rapid gains in frontier AI reasoning systems have positioned models close to the 90% threshold on FrontierMath, driving the strong 90% market-implied odds for a breakthrough before 2027.** FrontierMath, developed by Epoch AI, consists of hundreds of original, unpublished research-level mathematics problems across tiers of increasing difficulty, designed to resist contamination and test expert-level capabilities that previously required hours or days of human mathematician effort. As of mid-2026, top models have advanced dramatically: Claude Fable 5 (max) reached 87.8% on the updated Tier 4 (v2) benchmark released in June 2026, while OpenAI’s GPT-5.6 Sol variants have posted scores in the low-to-mid 80s on the same set and up to 89% on legacy snapshots. These results reflect iterative improvements in long-horizon reasoning, chain-of-thought techniques, and domain-specific training for mathematical domains such as number theory and algebraic geometry. Traders see continued momentum from the ongoing release cycle of advanced large language models (LLMs) and reasoning systems through late 2026, including potential GPT-6 iterations or refined Claude and Gemini releases. Historical precedent shows scores rising from under 2% at the benchmark’s 2024 launch to over 40% on easier tiers and near 90% on harder ones within roughly 18–20 months. Key swing factors include further scaling of inference-time compute, algorithmic refinements, and possible specialized “AI co-mathematician” training runs. While timelines can slip and exact benchmark versions matter for resolution, the current trajectory and proximity to the threshold underpin the high consensus that 90% will be achieved before the end of 2026.

This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No".

The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.
Обсяг
$117,315
Дата завершення
Dec 31, 2026
Ринок відкрито
Nov 12, 2025, 5:15 PM ET
This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.
This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used. **Rapid gains in frontier AI reasoning systems have positioned models close to the 90% threshold on FrontierMath, driving the strong 90% market-implied odds for a breakthrough before 2027.** FrontierMath, developed by Epoch AI, consists of hundreds of original, unpublished research-level mathematics problems across tiers of increasing difficulty, designed to resist contamination and test expert-level capabilities that previously required hours or days of human mathematician effort. As of mid-2026, top models have advanced dramatically: Claude Fable 5 (max) reached 87.8% on the updated Tier 4 (v2) benchmark released in June 2026, while OpenAI’s GPT-5.6 Sol variants have posted scores in the low-to-mid 80s on the same set and up to 89% on legacy snapshots. These results reflect iterative improvements in long-horizon reasoning, chain-of-thought techniques, and domain-specific training for mathematical domains such as number theory and algebraic geometry. Traders see continued momentum from the ongoing release cycle of advanced large language models (LLMs) and reasoning systems through late 2026, including potential GPT-6 iterations or refined Claude and Gemini releases. Historical precedent shows scores rising from under 2% at the benchmark’s 2024 launch to over 40% on easier tiers and near 90% on harder ones within roughly 18–20 months. Key swing factors include further scaling of inference-time compute, algorithmic refinements, and possible specialized “AI co-mathematician” training runs. While timelines can slip and exact benchmark versions matter for resolution, the current trajectory and proximity to the threshold underpin the high consensus that 90% will be achieved before the end of 2026.

This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No".

The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.
Обсяг
$117,315
Дата завершення
Dec 31, 2026
Ринок відкрито
Nov 12, 2025, 5:15 PM ET
This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.

Обережно з зовнішніми посиланнями.

Часті запитання

«AI model scores ≥ 90% on FrontierMath Benchmark before 2027?» — це ринок прогнозів на Polymarket, де трейдери купують і продають акції «Так» або «Ні» залежно від того, чи вірять вони, що ця подія станеться. Поточна краудсорсингова ймовірність — 90% для «Yes». Наприклад, якщо «Так» коштує 90¢, ринок колективно оцінює шанс цієї події в 90%. Ці шанси безперервно змінюються, коли трейдери реагують на нові події. Акції правильного результату погашаються по $1 кожна при вирішенні ринку.

Станом на сьогодні, «AI model scores ≥ 90% on FrontierMath Benchmark before 2027?» згенерував $117.3K загального обсягу торгів з моменту запуску ринку Nov 12, 2025. Цей рівень торгової активності відображає сильну залученість спільноти Polymarket та забезпечує, що поточні шанси базуються на глибокому пулі учасників ринку. Ви можете відстежувати рухи цін наживо та торгувати будь-яким результатом прямо на цій сторінці.

Щоб торгувати на «AI model scores ≥ 90% on FrontierMath Benchmark before 2027?», просто оберіть, чи вірите ви, що відповідь — «Так» або «Ні». Кожна сторона має поточну ціну, що відображає ймовірність ринку. Введіть суму та натисніть «Торгувати». Якщо ви купили акції «Так» і результат — «Так», кожна акція виплачує $1. Якщо «Ні» — ваші акції «Так» коштують $0. Ви також можете продати акції в будь-який час до вирішення.

Поточна ймовірність для «AI model scores ≥ 90% on FrontierMath Benchmark before 2027?» — 90% для «Yes». Це означає, що спільнота Polymarket вважає, що є 90% шанс, що ця подія станеться. Ці шанси оновлюються в реальному часі.

Правила вирішення для «AI model scores ≥ 90% on FrontierMath Benchmark before 2027?» точно визначають, що має статися для оголошення переможця — включаючи офіційні джерела даних. Ви можете переглянути повні критерії вирішення в розділі «Правила» на цій сторінці. Рекомендуємо уважно прочитати правила перед торгівлею.