# Current state
The market resolves on the LMArena (arena.ai) text leaderboard "Rank" at Dec 31, 2026, 12:00 PM ET, based on which company's model sits #1 (style control off). As of mid-August 2026, Anthropic's Claude Fable 5 (and/or Opus 5) reportedly holds or is tied for #1, but the frontier cluster (Anthropic, OpenAI, Google) is separated by only ~10-25 Elo points and leadership has rotated frequently over 2025-2026.
# Timeline of key events
- 2025-2026 (reported): Leadership on LMArena text leaderboard rotated multiple times among Google, OpenAI, Anthropic, xAI [claude_news].
- 2026-02: Claude Opus 4.6 / Opus 4.6 Thinking tied #1 at 1503 Elo; Gemini 3.1 Pro Preview #3 at 1500 [buildmvpfast.com, reported].
- 2026-02/03 (undated precisely): GPT-5.4 briefly overtook Opus 4.6, reaching 1502 Elo vs 1494 [mangomindbd.com, reported].
- 2026-07-01: Claude Fable 5 "restored" to #1 after reported removal/re-listing [localaimaster.com, reported].
- 2026-07-12: Score re-baseline pushes Fable 5 to ~1508-1525 Elo, #1 [localaimaster.com, reported].
- 2026-07-16: Bloomberg reports Google's Gemini 3.5 Pro is months behind schedule, no confirmed launch date [Bloomberg via claude_news, reported].
- 2026-07-21 / mid-Aug: Google ships Gemini 3.6 Flash and 3.7 Flash (efficiency-tier, not frontier flagship) [digitaltrends.com, confirmed release; strategic framing reported].
- 2026-07-24: Anthropic releases Claude Opus 5 (Anthropic's 4th Claude 5-family release in <2 months); Anthropic claims SOTA on Frontier-Bench/GDPval-AA coding/knowledge benchmarks [anthropic.com, axios.com, techcrunch.com — confirmed release, claims are company-sourced].
- 2026-08-13: Tracker (Artificial Analysis-style) shows Opus 5 #1, one point ahead of Fable 5; GPT-5.6 Sol and Grok 4.6 tied ~3% back [felloai.com, reported].
- 2026-08-12: Live arena.ai text leaderboard shows Elo range 952-1507 across 391 models, 7.78M votes; top model not named in snippet [arena.ai, confirmed data existence, model unconfirmed].
# Event
Will Anthropic (any Claude model) hold the #1 rank on the LMArena text leaderboard (style control off) as checked Dec 31, 2026, 12:00 PM ET.
# Outcomes to forecast
- Yes (Anthropic model ranks #1)
- No (any other company ranks #1)
# Kalshi market anchor
No kalshi_direct data was returned for this ticker. The only direct market data available is Polymarket for this exact ticker: **current YES price 68.5%**, up from a 54.00% low, near its 70.5% high; 7-day trend +2.0%, 30-day trend +3.0%; total volume ~$76,823 over 74 days. This is trending upward and should be treated as the primary consensus anchor in absence of Kalshi data.
# Sub-question answers
1. **Polymarket prices for group members** — No live multi-outcome group data was retrieved (polymarket_related found 0 matches). A code_execution tool produced only illustrative/placeholder figures (Anthropic ~21%, OpenAI 33%, Google 30%, xAI 8%, Meta 4%, DeepSeek 2%) explicitly caveated as NOT live data — unreliable, discard for calibration; use the single-market 68.5% YES as authoritative instead (note the sharp divergence between these two Anthropic estimates is a genuine gap).
2. **Current #1 holder & gap** — As of Aug 2026, reports converge that an Anthropic model (Claude Fable 5 or Opus 5) holds or ties #1 with Elo ~1507-1525, ahead of a tight cluster (Gemini 3.1 Pro Preview, GPT-5.5/5.6 Pro/Sol, Opus 4.8) within ~10-25 points [localaimaster.com, felloai.com — reported, not independently verified against live arena.ai].
3. **Historical Claude #1 status** — Claude models have held #1 multiple times in 2026 (Opus 4.6 in Feb, Fable 5 from July), but leadership swapped away at least once (GPT-5.4 briefly surpassed Opus 4.6) — indicating Claude's #1 tenure has been intermittent, not sustained for the full year [buildmvpfast.com, mangomindbd.com, localaimaster.com].
4. **Upcoming releases** — Anthropic has an unusually fast 2026 cadence (4 Claude-5-family releases in <2 months as of July: Mythos, Fable, Opus 4.7/4.8, Opus 5) [axios.com]. Google's Gemini 3.5 Pro (frontier flagship) is reportedly "months behind schedule" with no confirmed date, and Google has shipped only Flash-tier updates (3.6, 3.7) [Bloomberg via claude_news]. No specific GPT-6 or "Claude Opus 6" timeline found in research.
5. **Base rate of #1 persistence** — No verified empirical LMArena turnover count was found; a code_execution Fermi model (using assumed 3-9 leadership changes over 24 months) estimates P(current leader retains #1 at +12mo) ranges from ~26-35% (low-turnover assumption) down to ~25% (near-uniform, high-turnover assumption) — directionally suggests no lab has a strong structural lock on #1 over a 12-month horizon.
6. **Anthropic's submission practices to arena** — No direct evidence found that Anthropic deprioritizes LMArena; multiple sources describe an active July 2026 "restoration" and "re-baseline" event for Claude Fable 5, implying Anthropic (or the arena) actively manages/updates its listing, and Anthropic touts arena/benchmark performance in its own announcements [localaimaster.com, anthropic.com] — no evidence of systematic non-submission.
# Key facts (high-confidence, factual)
1. [polymarket_direct] This exact market's YES price is 68.5%, up from a 54% low, trending up over 7d/30d.
2. [anthropic.com/techcrunch] Anthropic released Claude Opus 5 on 2026-07-24, claiming SOTA on coding/knowledge benchmarks.
3. [Wikipedia/LMArena] LMArena is a public human-preference voting platform; OpenAI, Google DeepMind, and Anthropic all supply models to it.
4. [Bloomberg via claude_news] Google's next frontier model (Gemini 3.5 Pro) is reportedly delayed with no confirmed launch date as of mid-2026.
5. [arena.ai, confirmed] As of 2026-08-12 the live text leaderboard spans Elo 952-1507 across 391 models with 7.78M votes (top model name not captured in snippet).
# Cross-market signals
- Kalshi related: No directly comparable Kalshi market found (only an unrelated swimsuit-cover market matched keyword "best AI model").
- Polymarket: Same-ticker YES at 68.5%, uptrending; no separate multi-outcome Polymarket group ("Anthropic/Google/OpenAI/xAI…") was found live — any such breakdown cited elsewhere is illustrative, not sourced.
- Sportsbook implied: N/A.
# Analyst opinions and speculation
- Multiple aggregator/SEO sites (localaimaster, swfte, felloai, buildfastwithai) converge on Anthropic holding or tying #1 in Aug 2026, but these are noted as having inconsistent model names/Elo figures — treat as directionally suggestive, not precise [claude_news].
- Consensus view: "no universally best model" — frontier is a tightly clustered, rapidly rotating pack; Anthropic's edge is real but narrow and use-case dependent [buildfastwithai.com].
# Directional lean per outcome
- **Yes**: Supported by Anthropic's current (Aug 2026) #1/near-#1 position, aggressive release cadence, and Google's delayed frontier model. Opposed by: historically volatile leadership (Claude lost #1 to GPT-5.4 earlier in 2026), tight Elo gaps (10-25 points) making any release from OpenAI/Google/xAI capable of flipping rank, and 4.5 months remaining for competitors (esp. OpenAI GPT-6, delayed Gemini 3.5 Pro) to ship.
- **No**: Supported by base-rate turnover models suggesting ~65-75% chance leadership changes hands at least once over any 12-month window; Google's Gemini 3.5/4 Pro and OpenAI's next flagship remain wildcards for H2 2026.
# Gaps / unknowns
- No live Kalshi YES price was retrieved for this ticker (only Polymarket data for the same ID) — brief anchors on Polymarket 68.5%.
- No verified live arena.ai #1 model name as of the exact current date; only proxy/aggregator claims (Aug 2026 snapshots).
- No confirmed Gemini 3.5/4 Pro or GPT-6 launch dates before market close.
- Illustrative Polymarket multi-outcome breakdown (Anthropic 21%) is explicitly a placeholder, not live — creates unresolved tension with the 68.5% single-market price.
# Calibration anchors
- Polymarket (same ticker) current YES price: 68.5%, trending up (+2%/7d, +3%/30d).
- Fermi/base-rate model: ~25-35% chance a given non-uniform leader retains #1 after 12 months, depending on assumed turnover frequency (6-9+ changes/24mo).
- Historical precedent: Claude has held #1 intermittently through 2026 but lost it at least once to GPT-5.4, indicating no single lab has held sustained (>6mo) uncontested #1 status this year.