This market will resolve according to the third-highest ranked Company based on the arena.ai Text Arena (Math) when the table under the "Leaderboard" tab filtered for "Labs" is checked on September 30, 2026, 12:00 PM ET. Results from the "Lab Rank" column under the "Text Arena | Math" Leaderboard tab at https://arena.ai/leaderboard/text/math-no-style-control?rankBy=labs with style control off (Adjustments: None) and filtered for "Labs" will be used to resolve this market. AI companies will be ordered primarily by their Lab Rank at the market’s check time. If the results based on the lab ranking are ambiguous or unavailable, the relevant AI companies will be ordered according to their highest-ranking AI model in the leaderboard’s “Models” view. If two or more models are tied on rank, they will be ordered by their Arena score, including any underlying, unrounded, granular values reflected in the data below the leaderboard. If a tie still remains, alphabetical order of AI lab/company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact arena score, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies third place under this ranking. The resolution source for this market is the arena.ai Text Arena (Math). If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Alibaba holds the highest implied probability at 29% for third place among AI labs on the Text Arena Math leaderboard by September 30, driven by its recent Qwen3.8-Max release featuring 2.4 trillion parameters and targeted post-training for mathematical reasoning and long-horizon tasks. OpenAI, Google, Z.ai, and Moonshot cluster behind with lower odds, reflecting Anthropic's current dominance in crowd-sourced Elo ratings via models like Claude Fable 5 and Opus variants. Traders see tight competition among Chinese labs due to strong benchmark showings on AIME and HMMT math evaluations, yet arena outcomes hinge on real-time user votes that can shift quickly with model updates. Key upcoming catalysts include any final-week releases or fine-tunes before resolution, as small Elo gains could reorder lab standings in this closely contested category.