Trends
Related Markets
Related questions for context, not a comparison of equivalent contracts.
Next Google Gemini Pro Model: Humanity’s Last Exam Debut?
OpenAI’s Astra Model: Humanity’s Last Exam Debut?
Next Claude Opus Model: Humanity’s Last Exam Debut?
Highest score on Humanity's Last Exam before Dec 31, 2026?
Will Al achieve 85% or higher on the Humanity's Last Exam benchmark before 2027?
Will the first AI model that saturates Humanity's Last Exam be employable as a software engineer?
Rules
This market will resolve to "Yes" if the next Grok Model (4.6+) model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site.
+3 more outcomes
+5 more outcomes
+7 more outcomes
+4 more outcomes
+2 more outcomes