OpenAI’s GPT-6 Astra and GPT-5.4 Pro currently post Humanity’s Last Exam scores of 54.7–58.7 percent on leading September 2026 leaderboards, trailing Anthropic’s Claude Fable 5.1 and Opus 5 models at 59–65 percent. The 2,500-question benchmark, released in January 2025 by the Center for AI Safety and Scale AI, tests graduate-level reasoning across mathematics, physics, biology, and humanities with low contamination risk. Recent gains for OpenAI stem from iterative GPT-5 series releases and improved chain-of-thought scaffolding with tool use, yet Anthropic maintains an edge in adaptive multidisciplinary performance. Traders are watching for OpenAI’s next major model drop, potential Anthropic or Google advances, and any December 2026 releases that could push scores above 60 percent before year-end.
Polymarket ডেটা রেফারেন্স করে পরীক্ষামূলক AI-জেনারেটেড সারাংশ। এটি ট্রেডিং পরামর্শ নয় এবং এই মার্কেট কীভাবে রেজলভ হয় তাতে কোনো ভূমিকা রাখে না। · আপডেটেড$92,999 Vol.
55%+
84%
60%+
43%
65%+
24%
70%+
7%
$92,999 Vol.
55%+
84%
60%+
43%
65%+
24%
70%+
7%
For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
মার্কেট ওপেন হয়েছে: Jul 23, 2026, 6:53 PM ET
রেজলভার
0x65070BE91...For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
রেজলভার
0x65070BE91...OpenAI’s GPT-6 Astra and GPT-5.4 Pro currently post Humanity’s Last Exam scores of 54.7–58.7 percent on leading September 2026 leaderboards, trailing Anthropic’s Claude Fable 5.1 and Opus 5 models at 59–65 percent. The 2,500-question benchmark, released in January 2025 by the Center for AI Safety and Scale AI, tests graduate-level reasoning across mathematics, physics, biology, and humanities with low contamination risk. Recent gains for OpenAI stem from iterative GPT-5 series releases and improved chain-of-thought scaffolding with tool use, yet Anthropic maintains an edge in adaptive multidisciplinary performance. Traders are watching for OpenAI’s next major model drop, potential Anthropic or Google advances, and any December 2026 releases that could push scores above 60 percent before year-end.
Polymarket ডেটা রেফারেন্স করে পরীক্ষামূলক AI-জেনারেটেড সারাংশ। এটি ট্রেডিং পরামর্শ নয় এবং এই মার্কেট কীভাবে রেজলভ হয় তাতে কোনো ভূমিকা রাখে না। · আপডেটেড



বাহ্যিক লিংক থেকে সাবধান।
বাহ্যিক লিংক থেকে সাবধান।
সচরাচর জিজ্ঞাসা