This market will resolve according to the model that has the highest arena rank based on the arena.ai Text Arena (Overall) when the table under the "Leaderboard" tab is checked on the specified date, 12:00 PM ET. Results from the "Rank" column under the "Text Arena | Overall" Leaderboard tab at https://arena.ai/leaderboard/text/overall-no-style-control with style control off (Adjustments: None) and filtered for "Models" will be used to resolve this market. Note: Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score. No new model will be added to this market after market creation. Any model not explicitly listed in this market will be encompassed under the "Other" option. Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by their Arena score, including any underlying, unrounded, granular values reflected in the data below the leaderboard. If a tie still remains, alphabetical order of model names as listed in this market group (full string, including suffixes such as “-thinking”) will be used as a final tiebreaker (e.g., if two models remain tied, “claude-opus-4-6” would be ranked ahead of “claude-opus-4-6-thinking”). This market will resolve to the model that comes first according to this order. The resolution source for this market is the arena.ai Text Arena (Overall). If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Anthropic’s July 24 release of Claude Opus 5 has anchored trader sentiment, positioning the claude-opus-5-max variant at 90% implied probability for the best AI model by August 31. The model delivers step-change gains in deep reasoning, agentic coding, and long-horizon tool use while matching much of Claude Fable 5’s frontier intelligence at half the cost and the same $5/$25 per million token pricing as its predecessor. Official benchmarks highlight state-of-the-art results on Frontier-Bench and CursorBench at max effort, with improved efficiency and alignment, quickly establishing it as the default on Claude Max and strongest option on Pro. With no major competing releases in the ensuing weeks and historical patterns of rapid iteration favoring established leaders, market consensus reflects sustained dominance through the month-end resolution window.