# Current state
Anthropic's Claude Opus 5 currently sits at #1 on the arena.ai Code Arena | WebDev leaderboard (~1691–1703 Elo as of late Aug 2026), having succeeded a run of Anthropic Opus variants (4.5→4.6→4.7→4.8→5→Fable 5) that have held #1 nearly continuously since November 2025. The market resolves on a single leaderboard check on 2026-10-31, so today's rank is informative but not determinative — roughly 2 months remain for a new frontier release to unseat Anthropic.
# Timeline of key events
- 2025-11: Claude Opus 4.5 takes #1 on WebDev leaderboard, overtaking Gemini 3 Pro (top 3 within 20 Elo points). (confirmed — arena.ai/X post)
- 2026-04: Anthropic holds top 4 spots (Opus 4.6, Opus 4.6 Thinking, Sonnet 4.6, +1 more); GPT-5.2 first appears at #5, ~90 Elo behind. (reported — buildmvpfast.com)
- 2026-05: Claude Opus 4.7 Thinking leads; 4 of top 5 are Claude Opus variants; first OpenAI model outside top 10. (reported — propelcode.ai)
- ~2026-03: LMArena begins converting 10% of Direct-chat sessions into leaderboard-counted battles, tightening confidence intervals. (confirmed — arena.ai changelog)
- 2026-07: Kimi K3 (open-weight, 2.8T params) launches, cracking top 3 (~1674–1679 Elo), first open-weight model this high. (reported — LogRocket, GDELT)
- 2026-07-08: xAI ships Grok 4.5 (V9, 1.5T params) as an incremental coding model, not the flagship Grok 5. (reported — felloai.com)
- 2026-07-24: Claude Opus 5 released, becomes new WebDev Arena #1 (~1703–1704 Elo), dethroning "Fable 5." (reported — swfte.com, cryptobriefing.com)
- 2026-08: Arena.ai launches "Fullstack Code Arena"; OpenAI GPT-5.6 Sol scores ~1638, ~61 pts behind Anthropic's leader. (reported — cryptobriefing.com)
- 2026-08 (ongoing): Grok 5 still unreleased; xAI's flagship remains in training with no confirmed date, unlikely before Q3/Q4 2026. (reported — geotoolbox.ai, felloai.com)
# Event
Will Anthropic own the #1-ranked model on the arena.ai Code Arena | WebDev leaderboard when checked on 2026-10-31 12:00 PM ET?
# Outcomes to forecast
Yes / No
# Kalshi market anchor
No direct Kalshi price returned; using Polymarket sibling market (same event) as primary cross-market anchor: **current YES ~59.5%**, down 14pts over 7 days but up 8pts over 30 days, range 46.5%–76%, thin volume (~$34k total, 17 data points) — indicates active repricing and low liquidity, so noisy.
# Sub-question answers
1. **Current #1 holder** — Claude Opus 5 (Anthropic) leads at ~1691–1704 Elo; closest rival is open-weight Kimi K3 (~1674, ~29 Elo behind), then Grok 4.5/4.6 (~1630). [claude_news/LogRocket, ainexhub, cryptobriefing]
2. **Turnover over past 12 months** — Anthropic has held #1 essentially continuously since Nov 2025 (Opus 4.5 overtook Gemini 3 Pro), through successive Opus releases (4.6→4.7→4.8→5), i.e. ~10 of last ~10 months. Google briefly contested in Nov 2025 (within 20 Elo).
3. **Polymarket sibling prices** — Direct market data shows Anthropic YES at 59.5%. A separate normalized snapshot (code_execution, possibly stale) implies Anthropic ~51%, Google ~22%, OpenAI ~16%, xAI ~5%, Other ~7% after de-vig. Both show Anthropic as clear favorite but figures aren't fully reconciled (discrepancy: 59.5% vs 51%).
4. **Upcoming frontier releases** — Anthropic: Opus 5/Fable 5 already shipped; further updates plausible by Oct. Google: Gemini 3.x/3.5 Flash active but not topping WebDev board. OpenAI: GPT-5.5/5.6 (Sol, Terra, Codex variants) shipped but trailing 60+ Elo. xAI: Grok 5 still unreleased as of Aug 2026, unlikely before Q3/Q4 2026; only incremental Grok 4.5 shipped.
5. **Methodology changes** — WebDev leaderboard rebuilt under "Code Arena" branding; segmented into HTML and React sub-categories; Direct-chat votes now count toward battles (since ~March 2026), tightening CIs and stabilizing ranks faster — reduces noise-driven flips. [arena.ai changelog]
6. **Base rate for leader persistence** — Markov reversion model (uniform 25% floor across ~4 labs) gives Anthropic 26–40% persistence probability over a ~13-month horizon depending on assumed monthly turnover (p=0.15–0.35); over the shorter ~2-month remaining horizon to Oct 2026, persistence probability is much higher (roughly 70-85% under similar turnover assumptions), well above the naive long-horizon estimate.
# Key facts (high-confidence, factual)
1. [claude_news] Anthropic's Opus line has held #1 on WebDev Arena continuously since Nov 2025 through Aug 2026.
2. [claude_news] Nearest current rival is open-weight Kimi K3 (~29 Elo behind), not Google or OpenAI.
3. [claude_news] Grok 5 (xAI) unreleased as of Aug 2026; unlikely to launch and dominate before Oct 2026.
4. [arena.ai changelog] Leaderboard methodology recently changed (Direct-battle conversion, HTML/React segmentation) — increases rank stability.
5. [polymarket_direct] Market YES = 59.5%, volatile (46.5–76% range) on thin volume.
# Cross-market signals
- Kalshi related: not separately returned; treat Polymarket as primary anchor per tool hierarchy note.
- Polymarket: 59.5% YES, -14% (7d), +8% (30d); sibling-market de-vig estimate ~51% Anthropic (some inconsistency between snapshots).
- Sportsbook implied: N/A (not applicable to this event type).
# Analyst opinions and speculation
- Multiple blogs (LogRocket, PropelCode, Swfte) frame Anthropic's lead as durable, driven by Opus's coding/agentic specialization, with open-weight Kimi K3 as the main disruptive threat rather than Google/OpenAI.
- Code_execution's Markov model argues market may be over-pricing persistence versus naive base rates, but underweights Anthropic's apparent structural coding advantage.
# Directional lean per outcome
- **Yes (Anthropic)**: Strong — 10-month uninterrupted #1 streak, methodology now more stable, next real threat window (Grok 5) likely slips past Oct 2026, market prices 51–60%.
- **No (Other company)**: Risks include Kimi K3 (open-weight, closing gap to ~29 Elo) crossing #1, or a major Gemini/GPT release before Oct; base-rate reversion models suggest meaningful residual uncertainty over ~2 remaining months.
# Gaps / unknowns
- No official Kalshi-direct price was returned (only Polymarket) — true Kalshi consensus unconfirmed.
- Discrepancy between 59.5% (direct) and ~51% (normalized sibling) unresolved — possibly different timestamps.
- No confirmation of Anthropic's next release cadence between Aug–Oct 2026.
- "Other" resolution category (open-weight models, e.g., Kimi K3/Moonshot) not clearly mapped to a company outcome in this market's outcome set.
# Calibration anchors
- Polymarket YES (anchor): 59.5% (volatile, thin volume).
- Precedent: Anthropic has won essentially every WebDev leaderboard check since Nov 2025 (~10/10 months), a very strong recent base rate favoring continuation over a short 2-month window.