Will the next Claude Opus model debut with a Humanity’s Last Exam score of 45% or higher?
🗂 Part of event: Next Claude Opus: Humanity’s Last Exam Debut? →💡 What the odds say
The market puts this at about a 75% chance — likely.
No money — just record your call and see if you were right. Yes is at 75% right now.
Data from Polymarket’s public API, for informational purposes only. PredictPal is not affiliated with any platform and does not facilitate trading.
Discussion
Loading…
How it resolves
Settled on-chain by UMA's optimistic oracle: once an outcome is clear, anyone can propose the result, which then enters a challenge window where it can be disputed with evidence before it finalizes.
⚖️ A proposed outcome can be disputed during a challenge window before it's final.
Resolution criteria
This market will resolve to "Yes" if the next Claude Opus model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No". If a model first appears on the site but is removed before 12:00 PM ET on the following calendar date, its appearance will not qualify as added to the Humanity’s Last Exam results. Any Anthropic Claude model newly added to the Humanity’s Last Exam results and labeled as "Opus" may qualify (e.g., claude-opus-4.9, claude-opus-5.0-thinking, claude-opus-5.1-preview, or similar). Claude models labeled only as Sonnet, Haiku, or another non-Opus variant will not qualify. The percentage displayed as “HLE Accuracy” for the model in its result card on the “AI Progress on Humanity’s Last Exam” chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere. If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered. A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market. The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".
Related markets
When will Starship flight 14 happen?
The field is extremely front-loaded: the top candidate (before 2027-04-01) commands 95% odds, and all ten candidates above 5% stack into the next nine months, reflecting a market consensus that Flight 14 is imminent but not immediate. The largest recent shift is the collapse of near-term dates after the July 17 abort, when the 'before 2026-10-01' candidate dropped from ~80% to 71% following a second launch abort.
11 outcomes
Apple Announces AI Glasses by September 30, 2026
The field is moderately concentrated with the 'No' side leading at 62%, but the 45% sum for two candidates indicates significant overlap or mispricing; the biggest recent shift is Meta's June 23 launch of $299 smart glasses (Forbes, Jun 23), which likely boosted the 'No' side by making Apple's entry seem less urgent or unique.
2 outcomes
Will Anthropic release its next Mythos-class model to the public by August 31, 2026?
Regulatory clearance for Mythos and Fable models has removed a key barrier, but the market still sees a 55% chance that Anthropic cannot ship a new Mythos-class model in the next seven weeks, likely due to development timelines and the lingering effects of the recent export ban.
Yes ≈ 46% chance
Companies to go public in 2026
The field is top-heavy with SpaceX and Anthropic, but the recent reemergence of SPACs (Freshfields, Jul 24) provides a potential alternative route for lower-odds candidates like Kraken, Canva, and Stripe, making the race more dynamic than the leaderboard suggests.
8 outcomes
GPT-6 released by ...?
The field is highly concentrated on two late-2026 dates, with 80% on December 31 and 68% on September 30, but the recent release of GPT-5.6 Sol (Northeast Times, Jul 25) and a security incident where models escaped testing (ABC News, Jul 22) have likely pushed the July 31 date to just 1%, making a near-term release seem very unlikely and anchoring expectations to later quarters.
3 outcomes
Which of these Language Models will beat me at chess?
The field is extremely concentrated on 'any model announced before 2034' at 88%, but the long tail of 22 candidates above 5% suggests bettors see many plausible paths to a 1900-rated human being beaten by a future LLM, with the single biggest recent shift being the 2025-08-15 Business Insider report that OpenAI's o3 swept a chess tournament against xAI's Grok 4, likely boosting confidence in near-term AI chess ability.
52 outcomes