AI模型在2027年之前在FrontierMath Benchmark上的得分≥ 90% ?
商務AI科技大型科技公司AI基準
開始時間
2025-11-12
結束時間
2027-01-01
24小時成交量
$12K
總成交量
$134K
  • AI模型在2027年前於FrontierMath基準測試中得分≥90%?100¢

This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.

Rapid progress on FrontierMath, a benchmark of expert-level mathematics problems, drives the 99.7% market-implied probability that an AI model reaches ≥90% before 2027. As of early September 2026, OpenAI’s GPT-5.6 Sol leads the legacy version at 89% and FrontierMath v2 Tier 4 at 83%, following consistent gains from sub-2% at launch through GPT-5.5 releases and a June 2026 v2 update that corrected errors in 42% of problems and lifted scores across the board. Large language models continue to close the gap through scaled reasoning, tool use, and iterative releases, with trader consensus reflecting the historical pattern of math benchmarks saturating quickly once frontier systems exceed 50-70%. Remaining uncertainty centers on whether a new harder tier or version could reset the threshold before year-end.

AI模型在2027年之前在FrontierMath Benchmark上的得分≥ 90% ?

相關市場