Algum modelo de IA alcançará ___ Pontuação Geral da Arena até 30 de setembro?
IATecnologiaBig TechBenchmarks de IA
Data de início
2026-04-02
Data final
2026-10-01
Volume 24h
$19K
Volume total
$173K
  • Algum modelo de IA atingirá 1520 de Pontuação Geral de Arena até 30 de setembro de 2026?94¢
  • Algum modelo de IA atingirá 1530 de Pontuação Geral na Arena até 30 de setembro de 2026?
  • Algum modelo de IA alcançará uma Pontuação Geral de Arena de 1540 até 30 de setembro de 2026?
  • Algum modelo de IA atingirá 1550 de Pontuação Geral de Arena até 30 de setembro de 2026?

This market will resolve to "Yes" if any model on the Arena.AI Leaderboard (arena.ai/leaderboard/text) reaches at least the specified Arena Score by September 30, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". Results from the "Score" column under the "Text Arena | Overall" Leaderboard tab at https://lmarena.ai/leaderboard/text with style control off will be used to resolve this market. The resolution source for this market is the Chatbot Arena LLM Leaderboard found at arena.ai/leaderboard/text. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If permanently unavailable, this market will resolve to "No".

Top frontier large language models (LLMs) remain clustered between roughly 1505–1508 Arena Score on the LMSYS leaderboard as of early September 2026, led by Anthropic’s Claude Fable 5 and Opus 5 variants. Recent releases such as Google’s Gemini 3.8 Flash (High) at 1494 and Alibaba’s Qwen3.8-Max-0902, which topped specialized Code Arena categories, have delivered incremental gains rather than breakthroughs that would close the gap to higher thresholds. With just under four weeks until the September 30 resolution, trader consensus reflects the modest pace of post-training improvements and the absence of confirmed major model launches capable of adding the 10–15+ points needed. Competitive dynamics among Anthropic, Google, Meta, and Chinese labs continue to drive steady but contained progress, while historical patterns show frontier scores advancing only gradually outside major releases. Any surprise fine-tune or unannounced high-variance update could still shift odds in the final weeks, but the current leaderboard ceiling and timeline favor limited movement.

Algum modelo de IA alcançará ___ Pontuação Geral da Arena até 30 de setembro?

Mercados relacionados