Wird ein KI-Modell bis zum 30. September ___ Gesamt-Arena-Punktzahl erreichen?
AITechnikBig TechKI-Benchmarks
Startdatum
2026-04-02
Enddatum
2026-10-01
24h-Volumen
$19K
Gesamtvolumen
$173K
  • Wird ein KI-Modell bis zum 30. September 2026 einen Gesamt-Arena-Score von 1520 erreichen?94¢
  • Wird ein KI-Modell bis zum 30. September 2026 eine Gesamtarena-Punktzahl von 1530 erreichen?
  • Wird ein KI-Modell bis zum 30. September 2026 eine Gesamtpunktzahl von 1540 in der Arena erreichen?
  • Wird ein KI-Modell bis zum 30. September 2026 einen Gesamt-Arena-Score von 1550 erreichen?

This market will resolve to "Yes" if any model on the Arena.AI Leaderboard (arena.ai/leaderboard/text) reaches at least the specified Arena Score by September 30, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". Results from the "Score" column under the "Text Arena | Overall" Leaderboard tab at https://lmarena.ai/leaderboard/text with style control off will be used to resolve this market. The resolution source for this market is the Chatbot Arena LLM Leaderboard found at arena.ai/leaderboard/text. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If permanently unavailable, this market will resolve to "No".

Top frontier large language models (LLMs) remain clustered between roughly 1505–1508 Arena Score on the LMSYS leaderboard as of early September 2026, led by Anthropic’s Claude Fable 5 and Opus 5 variants. Recent releases such as Google’s Gemini 3.8 Flash (High) at 1494 and Alibaba’s Qwen3.8-Max-0902, which topped specialized Code Arena categories, have delivered incremental gains rather than breakthroughs that would close the gap to higher thresholds. With just under four weeks until the September 30 resolution, trader consensus reflects the modest pace of post-training improvements and the absence of confirmed major model launches capable of adding the 10–15+ points needed. Competitive dynamics among Anthropic, Google, Meta, and Chinese labs continue to drive steady but contained progress, while historical patterns show frontier scores advancing only gradually outside major releases. Any surprise fine-tune or unannounced high-variance update could still shift odds in the final weeks, but the current leaderboard ceiling and timeline favor limited movement.

Wird ein KI-Modell bis zum 30. September ___ Gesamt-Arena-Punktzahl erreichen?

Ähnliche Märkte