How the contract works
Probability
How the price has moved
Analysis
Context
What moves the probability
Style-uncontrolled scoring
The leaderboard version used for settlement has style control off, which historically rewards longer or more polished-sounding answers. Claude's typically more concise default style has sometimes cost it ground against rivals on this exact metric, a modest but persistent drag on Anthropic's odds.
Release timing before the check date
Arena rankings can shift within days of a new flagship model entering the voting pool. Whichever lab ships its strongest model closest to 31 December 2026 gets a fresh boost right when it matters most for settlement.
Competing flagship releases from Google and OpenAI
Gemini and GPT or o-series models have repeatedly held the top Arena rank over 2024 and 2025. A major release from either lab in the second half of 2026 is the single biggest threat to an Anthropic finish.
xAI as a third contender
Grok's rapid release cadence adds another lab capable of briefly topping the leaderboard. Its presence widens the field beyond a two-way Anthropic-versus-incumbents contest.
Tie-break rule favoring Anthropic alphabetically
If Anthropic and a rival finish with statistically indistinguishable unrounded scores, the alphabetical tiebreak favors a company starting with A over Google, OpenAI, or xAI. This is a small structural edge that only matters in a near-exact tie.
The case for
- Anthropic ships a flagship Claude update before the December snapshot that outperforms rivals in blind pairwise voting on the Arena's style-uncontrolled scoring.
- No rival lab releases a stronger model in the weeks immediately before 31 December 2026, when a late release could otherwise flip the top rank quickly.
- In a near-exact tie on unrounded Arena score, the alphabetical tiebreak rule works in Anthropic's favor.
- The market's current 67% level already reflects a view that Anthropic's recent releases have closed much of the historical gap to Google and OpenAI on this specific leaderboard.
The case against
- Google's Gemini and OpenAI's GPT or o-series models have repeatedly held the top Arena rank through 2024 and 2025, showing the top spot is genuinely contested rather than settled.
- The style-uncontrolled scoring format has historically been less favorable to Claude's typically concise responses than to more verbose competitor outputs.
- xAI's Grok adds a third serious contender capable of topping the leaderboard on short notice after a strong release.
- Trading volume of $62,772 on a single venue is thin, meaning the current level may not have absorbed much information about labs' late-2026 release plans.
What to watch
Trade this contract
- gas covered
