¿Algún modelo de IA alcanzará ___ Puntuación general de Arena antes del 31 de diciembre?
IAgoogleTechBig TechPuntos de referencia de IA
Fecha de inicio
2026-01-02
Fecha de fin
2027-01-01
Volumen 24h
$11K
Volumen total
$166K
  • ¿Algún modelo de IA alcanzará una puntuación de al menos 1550 en Chatbot Arena antes del 31 de diciembre?14¢
  • ¿Algún modelo de IA alcanzará una puntuación de al menos 1600 en el Chatbot Arena antes del 31 de diciembre?
  • ¿Algún modelo de IA alcanzará una puntuación de al menos 1650 en Chatbot Arena antes del 31 de diciembre?
  • ¿Algún modelo de IA alcanzará una puntuación de al menos 1700 en Chatbot Arena antes del 31 de diciembre?
  • ¿Algún modelo de IA alcanzará una puntuación de al menos 1530 en el Chatbot Arena antes del 31 de diciembre?50¢
  • ¿Algún modelo de IA alcanzará una puntuación de al menos 1520 en Chatbot Arena antes del 31 de diciembre?96¢
  • ¿Algún modelo de IA alcanzará una puntuación de al menos 1540 en la Chatbot Arena antes del 31 de diciembre?32¢

This market will resolve to "Yes" if any model on the Chatbot Arena LLM Leaderboard (https://lmarena.ai/) reaches at least the specified Arena Score by December 31, 2026, 11:59 PM ET. Otherwise, this market will resolve to "No". Results from the 'Score' section on the 'Text Arena' Leaderboard tab (https://lmarena.ai/leaderboard/text), with the style control unchecked, will be used to resolve this market. The resolution source is the Chatbot Arena LLM Leaderboard (https://lmarena.ai/). If this source is temporarily unavailable, the market remains open until it is accessible again; if permanently unavailable, this market will resolve to "No".

Recent model releases from Anthropic, Google, and Alibaba continue to drive gains on the LMSYS Arena leaderboard, where top overall Elo scores hover near 1505-1508 as of early September 2026. Claude Fable 5 and Opus variants lead the text arena, while fresh entries like Gemini 3.8 Flash (1494 pts) and Qwen3.8-Max-0902 demonstrate rapid iteration in coding and agentic tasks. Competitive pressure among frontier labs, including Meta's Muse Spark updates and open-weight challengers from Moonshot and Zhipu, fuels frequent capability jumps. With roughly four months remaining until year-end, further post-training refinements, larger context windows, and specialized variants could push scores higher, though historical patterns show incremental rather than explosive monthly gains. Traders should monitor upcoming developer conferences and lab announcements for catalysts that could shift the threshold.

¿Algún modelo de IA alcanzará ___ Puntuación general de Arena antes del 31 de diciembre?

Mercados relacionados