2027년 이전 FrontierMath 벤치마크에서 AI 모델 점수가 90% 이상입니까?
비즈니스AI기술빅 테크AI 벤치마크
시작일
2025-11-12
종료일
2027-01-01
24시간 거래량
$12K
총 거래량
$134K
  • AI 모델이 2027년 이전에 FrontierMath 벤치마크에서 90% 이상 점수를 받을까요?100¢

This market will resolve to "Yes" if a state-of-the-art (SOTA) AI model achieves a score of 90% or greater on the FrontierMath Exam by December 31, 2026, 11:59 PM ET. Otherwise, the market will resolve to "No". The primary resolution source will be information from EpochAI however a consensus of credible reporting may also be used.

Rapid progress on FrontierMath, a benchmark of expert-level mathematics problems, drives the 99.7% market-implied probability that an AI model reaches ≥90% before 2027. As of early September 2026, OpenAI’s GPT-5.6 Sol leads the legacy version at 89% and FrontierMath v2 Tier 4 at 83%, following consistent gains from sub-2% at launch through GPT-5.5 releases and a June 2026 v2 update that corrected errors in 42% of problems and lifted scores across the board. Large language models continue to close the gap through scaled reasoning, tool use, and iterative releases, with trader consensus reflecting the historical pattern of math benchmarks saturating quickly once frontier systems exceed 50-70%. Remaining uncertainty centers on whether a new harder tier or version could reset the threshold before year-end.

2027년 이전 FrontierMath 벤치마크에서 AI 모델 점수가 90% 이상입니까?

관련 시장