# Current state
As of early-to-mid August 2026, Anthropic's Claude Fable 5 is reported to hold #1 on the Arena.ai Text Arena (Overall) leaderboard (~1508-1525 Elo), per multiple August 2026 blog/tracker snapshots (felloai.com, localaimaster.com). However, the top tier (Fable 5, Opus 4.8, GPT-5.5 Pro, Gemini 3.1 Pro Preview) is described as extremely tight (~20-30 Elo points) and volatile, and this market resolves on a snapshot check on 2026-09-30, ~7 weeks after the cited data.
# Timeline of key events
- 2026-02: Claude Opus 4.6 reportedly becomes first model to hold #1 simultaneously on LMArena text, code, and search boards (reported, buildmvpfast.com).
- 2026-05: Claude Opus 4.6 holds ~1504 Elo, statistically tied with Gemini 3.1 Pro Preview (reported, productleadersdayindia.org).
- 2026-06 (June): Claude Fable 5 launches at #1 across Agent/Text/Code Arena (reported, Arena.ai official X post).
- 2026-06 (~mid): Fable 5 removed from Arena following US export-control directive suspending Anthropic access (confirmed via Arena.ai X post).
- 2026-07-01: Fable 5 restored to Arena (reported, Arena.ai X post).
- 2026-07-09: OpenAI ships GPT-5.6 (Sol/Terra/Luna) as new flagship (confirmed, Wikipedia/en.wikipedia.org/wiki/GPT-5.6).
- 2026-07-09: xAI ships Grok 4.5 (current flagship pre-Grok 5) (reported, nxcode.io).
- 2026-07-12: Fable 5 Elo re-baselined, regains/confirms #1 (~1525 Elo) (reported, localaimaster.com).
- 2026-07-19/20: Alibaba previews Qwen3.8, claims #2 behind Claude Fable 5 (reported, siliconangle.com/chinatechnews.com).
- 2026-07-21: Google reportedly ships three smaller/cheaper Gemini models instead of flagship Gemini 3.5 Pro, which remains delayed with no new timeline (reported, Reuters via tech-insider.org).
- 2026-07-24: Anthropic launches Claude Opus 5 flagship (confirmed, Axios) — ranks #6 on general text (~1491.8) but #1 on WebDev, image-to-WebDev, and Agent boards (reported, felloai.com).
- 2026-07-26: Moonshot's Kimi K3 open-weights released; takes #1 on Frontend Code Arena but only ~#6 on general text (reported, swfte.com).
- 2026-08-01 (vote cutoff cited): Claude Fable 5 #1 on Text Arena Overall (1508.6 Elo), also #1 on HLE and AA-Omniscience (reported, felloai.com).
- 2026-08-06: X account reports Gemini 3.5 Pro spotted in secret Arena preview, suggesting imminent launch (rumored, @teortaxesTex via orcarouter.ai).
# Event
Resolves YES if Anthropic's model holds Arena.ai Text Arena (Overall, no style control) Rank #1 at the 2026-09-30 12:00 PM ET check.
# Outcomes to forecast
Yes, No
# Kalshi market anchor
No kalshi_direct price was returned in this research pull (only kalshi_related sibling markets, e.g., "OpenAI or Anthropic IPO first — Anthropic" at 83%, unrelated to model leaderboard). **Treat this as a gap**: use the sibling Polymarket price on the identical ticker/question as the best available cross-market anchor.
# Sub-question answers
1. **Current Rank #1 and margin** — Multiple August 2026 trackers report Claude Fable 5 (Anthropic) at #1, ~1508-1525 Elo, ahead of a tight cluster (Claude Opus 4.8, GPT-5.5 Pro, Gemini 3.1 Pro Preview) within ~20-30 Elo points [felloai.com, localaimaster.com]. No single authoritative live pull of arena.ai was available; figures come from third-party blogs, not the primary source.
2. **Anthropic's historical #1 status** — Yes, repeatedly: Opus 4.6 held #1 (and swept text/code/search) in Feb 2026; Fable 5 held #1 from June 2026 launch (with a temporary export-control removal), restored July 2026 [buildmvpfast.com, Arena.ai X].
3. **Frequency of #1 changing hands** — Not directly quantified by primary sources; qualitative evidence describes a "tight," "volatile," multi-way race with different leaderboards (LMArena vs. Artificial Analysis) crowning different leaders in the same period (e.g., Gemini 3 Pro vs GPT-5.2 vs Fable 5 across Jan–Aug 2026) [felloai.com]. The code_execution tool's "4 of 18 months" base rate is explicitly flagged as an illustrative/placeholder estimate, not real data — unreliable.
4. **Upcoming frontier releases** — Anthropic: Opus 5 shipped July 24, 2026 (4th Claude-5-series release in <2 months); further releases plausible before Sept 30. Google: Gemini 3.5 Pro repeatedly delayed (missed June, July targets), possibly imminent per an Aug 6 leak. OpenAI: GPT-5.6 shipped July 9; GPT-6 unconfirmed, Polymarket-implied ~71% by Sept 30, 2026 (per a cited secondary source, not verified live). xAI: Grok 5 unreleased, no confirmed date; Grok 4.5 (July 9) is current flagship [Axios, Wikipedia, tech-insider.org, nxcode.io].
5. **Sibling Polymarket normalization** — Only this market's own Polymarket price (81.5% Yes for Anthropic) was retrieved; no live prices for Google/OpenAI/xAI sibling outcomes were found (polymarket_related returned 0 matches). The code_execution "de-vigged" breakdown (OpenAI 37.6%, Google 32.7%, Anthropic 18.8%) is explicitly a fabricated placeholder and should be disregarded — it contradicts the real Polymarket 81.5% price for Anthropic on this exact ticker.
6. **Anthropic's optimization target** — Evidence suggests Anthropic optimizes broadly: Opus 5 dominates coding/agent boards (#1 WebDev, Agent) but ranks lower (#6) on general Text Arena, while Fable 5 (a different, "most powerful" model) is Anthropic's Text-Arena-optimized flagship holding #1 [felloai.com]. This implies Anthropic runs parallel flagships — one for agentic/coding evals, one competitive on arena-style preference — rather than sacrificing arena rank entirely.
# Key facts (high-confidence, factual)
1. [polymarket_direct] This exact market's Polymarket price: 81.5% Yes, down from a 30-day high of 89.5% (7d: -7pp, 30d: -4.5pp), volume ~$24.3K.
2. [Axios] Anthropic released Claude Opus 5 on 2026-07-24.
3. [Wikipedia/GPT-5.6] OpenAI released GPT-5.6 on 2026-07-09.
4. [Arena.ai X] Claude Fable 5 was removed from Arena mid-2026 due to a US export-control directive, then restored 2026-07-01.
5. [Wikipedia/Claude] US federal agencies moved to restrict Claude use over weapons/surveillance policy disputes (DoD "supply chain risk" designation, injunction 2026-03-26) — a nonmarket political risk factor for Anthropic's federal footprint (not directly resolution-relevant but signals regulatory friction).
# Cross-market signals
- Kalshi related: No direct arena/model-quality Kalshi market found; only tangential Anthropic markets (IPO race 83%, sector classification 85%) — not informative for this question.
- Polymarket (own market): 81.5% Yes, trending down (-7pp/7d, -4.5pp/30d) — suggests eroding but still strong confidence in Anthropic.
- Polymarket sibling outcomes (Google/OpenAI/xAI/Other): Not found live; any numeric breakdown circulated should be treated as unverified/fabricated.
# Analyst opinions and speculation
- claude_news synthesis: "Anthropic's lead is real but far from secure through year-end," citing imminent Gemini 3.5 Pro, uncertain GPT-6, and Grok 5 as wildcards.
- Third-party trackers disagree on ranking specifics (Elo values differ by ~15-20 points across sources), reflecting methodological noise in unofficial arena.ai mirrors rather than the primary site.
# Directional lean per outcome
- **Yes (Anthropic)**: Supported by consistent, repeated #1 claims (Opus 4.6 Feb, Fable 5 June–Aug) across multiple trackers, plus the real Polymarket price at 81.5%. Opposed by: 7-week gap to resolution, historically tight/volatile top tier, imminent Gemini 3.5 Pro, possible GPT-6, and a documented pattern of #1 changing hands across labs within the same year.
- **No (other company)**: Supported by declining trend on Polymarket (-7pp/7d), Gemini 3.5 Pro spotted in Arena preview (Aug 6), rapid multi-lab release cadence increasing chance of overtake by Sept 30. Opposed by lack of any confirmed rival #1 claim as of the most recent (August) snapshot.
# Gaps / unknowns
- No live/authoritative arena.ai pull was obtained; all Elo/rank data are third-party paraphrases with inconsistent numbers.
- No actual Kalshi YES price for this specific ticker was retrieved (kalshi_direct tool output absent).
- No verified Polymarket sibling-outcome prices for Google/OpenAI/xAI/Other; the only numeric breakdown offered was explicitly a fabricated placeholder.
- No quantified historical base rate of #1-leader turnover exists in reliable sourcing (only an unreliable placeholder estimate).
# Calibration anchors
- Polymarket price for this exact market (best available cross-market anchor): 81.5% Yes, trending down from 89.5% over 30 days.
- Qualitative precedent: Anthropic has held #1 in at least 3 distinct windows in 2026 (Feb, May, June–Aug), but leadership has also been contested/tied repeatedly by Gemini and GPT variants — suggests moderate-high but not overwhelming confidence, consistent with market pricing in the high-70s to low-80s range rather than90%+.