TOTAL VOLUME:

$134b

24H VOL:

$103,397,351

24H TRANSACTIONS:

2,388,728,490

OPEN INTEREST:

$1,410,176,180

399,592

Markets across

30,097

events

MATCHED EVENTS:

2,622

PLATFORM COVERAGE:

5

Polymarket:

39%

VS.

Kalshi:

61%

BETA
Dashboards
Tour
All
Other
Next Mythos-Class Model: Humanity's Last Exam Debut?
polymarket

Next Mythos-Class Model: Humanity's Last Exam Debut?

Volume:
$23,059

50%+

 - Polymarket

50%+ - Polymarket

1W

News

Positive

Negative

Neutral

Hover marker for details

Vol.

·

Resolved Sep 6, 2026

Closed: Sep 6, 2:01 PM EST

polymarket

Polymarket

View
Outcome
Trade
Chance %
Price
Spread
Liquidity
Volume
24h
7d
Open Interest
Ends in
Result
polymarket

50%+

View
100%
1%
Yes 100¢No 0¢
0.1¢
N/A
$8,827
N/A
N/A
N/A
3mo 3d
Yes
polymarket

45%+

100%
2%
Yes 100¢No 0¢
0.5¢
N/A
$3,151
N/A
N/A
N/A
3mo 3d
Yes
polymarket

60%+

0%
2%
Yes 0¢No 100¢
—
N/A
$5,927
N/A
N/A
N/A
3mo 3d
No
polymarket

55%+

0%
4%
Yes 0¢No 100¢
0.8¢
N/A
$5,155
N/A
N/A
N/A
3mo 3d
No
Total markets: 4

Description

This market will resolve to "Yes" if the next Anthropic Mythos-class model added to the Humanity's Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No". Any model whose official name includes "Mythos," or "Fable" (e.g., Claude Fable 5.1 or Claude Fable 6), or that Anthropic officially describes as a "Mythos-class" model or similar, will qualify for this market's resolution. Anthropic's other model tiers, such as Opus, Sonnet, or Haiku, will not qualify. The percentage displayed as "HLE Accuracy" for the model in its result card on the "AI Progress on Humanity's Last Exam" chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model's Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere. If multiple qualifying models are added to the Humanity's Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model's first appearance will not be considered. If a model is first added to the Humanity's Last Exam results but is subsequently removed and not re-added, such that it is not displayed on the site at 12:00 PM ET on the calendar date following its first appearance, its appearance will not qualify as added to the Humanity's Last Exam results. A qualifying model must be newly added to the Humanity's Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market. The resolution source for this market is the official Humanity's Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model's HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".

Polymarket

This market will resolve to "Yes" if the next Anthropic Mythos-class model added to the Humanity's Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No". Any model whose official name includes "Mythos," or "Fable" (e.g., Claude Fable 5.1 or Claude Fable 6), or that Anthropic officially describes as a "Mythos-class" model or similar, will qualify for this market's resolution. Anthropic's other model tiers, such as Opus, Sonnet, or Haiku, will not qualify. The percentage displayed as "HLE Accuracy" for the model in its result card on the "AI Progress on Humanity's Last Exam" chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model's Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere. If multiple qualifying models are added to the Humanity's Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model's first appearance will not be considered. If a model is first added to the Humanity's Last Exam results but is subsequently removed and not re-added, such that it is not displayed on the site at 12:00 PM ET on the calendar date following its first appearance, its appearance will not qualify as added to the Humanity's Last Exam results. A qualifying model must be newly added to the Humanity's Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market. The resolution source for this market is the official Humanity's Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model's HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".

Frequently asked questions

Currently, analysts broadly anticipate the next Mythos-Class model will debut with a Humanity's Last Exam score somewhere between 50% and 60%. This market reflects a strong consensus among traders that the score will exceed 55%, with the implied probability of that outcome being very high. While analyst forecasts offer expert opinions, this market aggregates the wisdom of the crowd, providing a real-time assessment of probabilities based on actual financial commitments. Discrepancies between the two can highlight areas of disagreement or uncertainty surrounding the model’s capabilities.

On Polymarket, this market is priced through a continuous double auction, where traders buy and sell shares representing their belief in the outcome. The price of a share directly reflects the probability of that outcome occurring, as determined by supply and demand. On Polymarket, prices reflect that venue's order book, liquidity, and how traders price the outcome right now. Traders can adjust their positions as new information becomes available, influencing the market price and overall probability assessment. The current price indicates a high degree of confidence in the specified outcome, reflecting substantial trading activity and a strong prevailing sentiment.

This market resolves around Dec 31, 2026, with the outcome confirmed once the event is verifiable from credible public reporting. Specifically, the resolution will depend on the publicly released Humanity's Last Exam score for the next Mythos-Class model. The score will be verified against credible public sources to determine whether it meets or exceeds the threshold specified in the market’s outcome conditions. Traders should monitor official announcements from the developers of the model for the final score and resolution details.

Several signals could significantly influence this market. Any announcements regarding the development timeline of the next Mythos-Class model, including potential delays or advancements, would likely cause price fluctuations. Furthermore, leaks or early previews of the model’s performance on benchmark tests, particularly those related to the Humanity's Last Exam, could shift trader sentiment. Positive or negative press coverage surrounding the model’s capabilities or the company developing it could also impact trading activity and the overall probability assessment reflected in this market.