This market will resolve according to the third-highest ranked Company based on the arena.ai Agent Arena Leaderboard when the table under "Agent Arena" filtered for "Labs" is checked on October 31, 2026, 12:00 PM ET. Results from the "Rank" column under the "Agent Arena" Leaderboard tab at https://arena.ai/leaderboard/agent filtered for "Labs" will be used to resolve this market. Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score. AI companies will be ordered primarily by their Lab Rank at the market’s check time. If the results based on the lab ranking are ambiguous or unavailable, the relevant AI companies will be ordered according to their highest-ranking AI model in the leaderboard’s “Models” view. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of AI lab/company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies third place under this ranking. The resolution source for this market is the arena.ai Agent Arena Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Moonshot leads trader sentiment for third-best AI agent lab by end-October at 29.5% implied probability, narrowly ahead of OpenAI at 22.5%, as recent agentic benchmarks highlight its July 2026 Kimi K3 release—a 2.8-trillion-parameter MoE with strong tool-use, long-context reasoning, and cost efficiency that narrows gaps with closed US models. OpenAI's GPT-5.6 variants show solid performance on Terminal-Bench and coding agents, while Anthropic's Claude Opus 5 and Mythos 5 dominate several indexes but trade lower here amid a crowded field. Competitive dynamics hinge on rapid iteration in multimodal agents, open-weight accessibility versus proprietary safety features, and developer adoption metrics, with Alibaba's Qwen and others adding pressure; new benchmarks or releases before October could readily shift the close consensus.