TOTAL VOLUME:
$134b
24H VOL:
$107,351,958
24H TRANSACTIONS:
2,388,728,490
OPEN INTEREST:
$1,416,970,024
400,720
Markets across
30,097
events
MATCHED EVENTS:
2,633
PLATFORM COVERAGE:
5
Polymarket:
39%
VS.
Kalshi:
61%
Closed: Jul 11, 7:59 PM EST
Polymarket
This market will resolve according to the model that has the highest arena rank based on the Chatbot Arena LLM Leaderboard (https://lmarena.ai/) when the table under the "Leaderboard" tab is checked on the specified date, 12:00 PM ET. Results from the "Rank" column under the "Text Arena | Overall" Leaderboard tab at https://arena.ai/leaderboard/text/overall-no-style-control with style control off will be used to resolve this market. No new model will be added to this market after market creation. Any model not explicitly listed in this market will be encompassed under the "Other" option. Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by their Arena score, including any underlying, unrounded, granular values reflected in the data below the leaderboard. If a tie still remains, alphabetical order of model names as listed in this market group (full string, including suffixes such as “-thinking”) will be used as a final tiebreaker (e.g., if two models remain tied, “claude-opus-4-6” would be ranked ahead of “claude-opus-4-6-thinking”). This market will resolve to the model that comes first according to this order. The resolution source for this market is the Chatbot Arena LLM Leaderboard found at https://lmarena.ai/. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve based on another resolution source.
This market will resolve according to the model that has the highest arena rank based on the Chatbot Arena LLM Leaderboard (https://lmarena.ai/) when the table under the "Leaderboard" tab is checked on the specified date, 12:00 PM ET. Results from the "Rank" column under the "Text Arena | Overall" Leaderboard tab at https://arena.ai/leaderboard/text/overall-no-style-control with style control off will be used to resolve this market. No new model will be added to this market after market creation. Any model not explicitly listed in this market will be encompassed under the "Other" option. Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by their Arena score, including any underlying, unrounded, granular values reflected in the data below the leaderboard. If a tie still remains, alphabetical order of model names as listed in this market group (full string, including suffixes such as “-thinking”) will be used as a final tiebreaker (e.g., if two models remain tied, “claude-opus-4-6” would be ranked ahead of “claude-opus-4-6-thinking”). This market will resolve to the model that comes first according to this order. The resolution source for this market is the Chatbot Arena LLM Leaderboard found at https://lmarena.ai/. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve based on another resolution source.
Prediction market odds often diverge from traditional analyst forecasts because they incorporate real-time financial incentives and crowdsourced information. While academic researchers and industry analysts may publish periodic evaluations of AI model performance, this market aggregates the views of traders who have money at stake. Analysts typically release reports on specific dates, whereas prediction markets update continuously as new benchmarks, research papers, or model releases emerge. The market price reflects a dynamic consensus that can shift rapidly, sometimes leading or lagging behind published expert opinions depending on information asymmetries and market efficiency.
On Polymarket, traders set prices through an automated market maker mechanism where each outcome token trades against a liquidity pool. On Polymarket, prices reflect that venue's order book, liquidity, and how traders price the outcome right now. The current odds reflect the probability that a particular AI model will be deemed the best on the resolution date, with prices ranging from near-zero to near-certain. Buyers and sellers continuously adjust their positions, and the market price converges toward an equilibrium that balances supply and demand. Higher prices indicate stronger market confidence in that outcome, while lower prices suggest skepticism or lower perceived probability among active traders.
This market resolves around Jul 11, 2026, at which point the outcome is confirmed based on verifiable information from credible public sources. The determination of which AI model is best will be assessed according to established benchmarks, performance metrics, and industry consensus at that time. Once the event is confirmed and documented, the market settles and traders receive payouts proportional to their positions. The resolution process is designed to be objective and transparent, ensuring that all participants can independently verify the outcome.
Major catalysts include new AI model releases, published benchmark results, and performance comparisons from reputable research institutions. Academic papers demonstrating significant capability improvements or novel architectures could shift trader sentiment substantially. Real-world deployment successes or failures, regulatory announcements affecting AI development, and competitive announcements from leading labs may also influence prices. Additionally, unexpected breakthroughs in reasoning, multimodal capabilities, or efficiency could reshape market expectations. Traders monitor industry conferences, preprint servers, and official model announcements closely, as any credible evidence of relative model performance typically triggers rapid repricing.