Will the next Claude Opus model debut with a Humanity’s Last Exam score of 55% or higher?
🗂 Part of event: Next Claude Opus: Humanity’s Last Exam Debut? →💡 What the odds say
The market puts this at about a 22% chance — unlikely.
No money — just record your call and see if you were right. Yes is at 22% right now.
Data from Polymarket’s public API, for informational purposes only. PredictPal is not affiliated with any platform and does not facilitate trading.
Discussion
Loading…
How it resolves
Settled on-chain by UMA's optimistic oracle: once an outcome is clear, anyone can propose the result, which then enters a challenge window where it can be disputed with evidence before it finalizes.
⚖️ A proposed outcome can be disputed during a challenge window before it's final.
Resolution criteria
This market will resolve to "Yes" if the next Claude Opus model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No". If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results. Any Anthropic Claude model newly added to the Humanity’s Last Exam results and labeled as "Opus" may qualify (e.g., claude-opus-4.9, claude-opus-5.0-thinking, claude-opus-5.1-preview, or similar). Claude models labeled only as Sonnet, Haiku, or another non-Opus variant will not qualify. The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere. If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered. A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market. The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".
Related markets
When will Starship flight 14 happen?
The field is extremely front-loaded: the top candidate (before 2027-04-01) commands 95% odds, and all ten candidates above 5% stack into the next nine months, reflecting a market consensus that Flight 14 is imminent but not immediate. The largest recent shift is the collapse of near-term dates after the July 17 abort, when the 'before 2026-10-01' candidate dropped from ~80% to 71% following a second launch abort.
11 outcomes
Apple Announces AI Glasses by September 30, 2026
2 outcomes
Which of these Language Models will beat me at chess?
The field is extremely concentrated on 'any model announced before 2034' at 88%, but the long tail of 22 candidates above 5% suggests bettors see many plausible paths to a 1900-rated human being beaten by a future LLM, with the single biggest recent shift being the 2025-08-15 Business Insider report that OpenAI's o3 swept a chess tournament against xAI's Grok 4, likely boosting confidence in near-term AI chess ability.
52 outcomes
Will Anthropic release its next Mythos-class model to the public by August 31, 2026?
Yes ≈ 46% chance