# Current state
The market resolves off the arena.ai Text Arena (Math) "Lab Rank" leaderboard snapshot taken 2026-08-31 12:00 ET. As of research date, no direct confirmation of arena.ai's current Lab Rank ordering was retrieved (tool couldn't access it directly); evidence instead comes from proxy math benchmarks (FrontierMath, KEAR AI Math Arena, general math leaderboards) showing a fluid three-way OpenAI/Google/Anthropic contest, with OpenAI most often #1 as of mid-2026 and Google/Anthropic contesting #2.
# Timeline of key events
- 2025 (mid): Google DeepMind's AlphaProof/AlphaGeometry 2 achieve silver-medal standard at IMO — confirmed (deepmind.google).
- 2025 (IMO): Google DeepMind and OpenAI both score 35/42 (gold-level); Google's result IMO-coordinator-certified, OpenAI's self-evaluated only — confirmed (gizmodo.com, intuitionlabs.ai).
- 2026-02: Gemini 3 Pro/Deep Think takes #1 on KEAR AI Math Arena, ending OpenAI's reign; GPT-5.2 High drops to #4 tied with Claude Opus 4.5; Moonshot (Chinese startup) reaches podium — reported (kearai.com).
- 2026-06: FrontierMath v2 shows GPT-5.5 Pro (87.7%) vs Claude Fable 5 (87.0%) essentially tied at top; Google not cited as leading this benchmark — reported (digitalapplied.com).
- 2026-05 to 07: Primary FrontierMath leaderboard shows OpenAI's GPT-5.6 Sol leading at ~89-89.0%, ahead of other OpenAI variants; Google not in top spots listed — reported (llm-stats.com, benchlm.ai).
- 2026-08-02: General "Mathematics" leaderboard shows GPT-5.2 Pro #1 (99.0%), GPT-5 Codex #2 (98.7%), DeepSeek V3.2 Speciale #3 (96.7%) — Google absent from top 3 — reported (pricepertoken.com).
- Polymarket price history: over past 30 days ranged 19%-72%, with a sharp -41.5% drop in the last 7 days to current 28.5% — confirmed via polymarket_direct.
# Event
Will Google DeepMind hold the #2 Lab Rank spot on the arena.ai Text Arena (Math) leaderboard at the Aug 31, 2026 12:00 ET check?
# Outcomes to forecast
Yes / No
# Kalshi market anchor
No direct Kalshi price was returned for this ticker in the research (kalshi_direct data absent from raw research); only Polymarket data available. Treat Polymarket price as best available consensus proxy: **28.5% YES**, down sharply from a 7-day-ago level implying ~70% (7d change -41.5pp), but up slightly over 30 days (+6pp). Volume is thin ($17.9k total, 22 data points) — low liquidity, high noise risk.
# Sub-question answers
1. **Current Polymarket price?** — 28.5% YES as of latest snapshot; extremely volatile (range 19%-72% over 30 days), suggesting large swings tied to individual model releases/benchmark news (polymarket_direct).
2. **Recent news affecting Google's math AI standing?** — Mixed: Google (Gemini 3 Pro/Deep Think) briefly led KEAR AI Math Arena in Feb 2026, but by mid-to-late 2026 OpenAI (GPT-5.5/5.6 series) reclaimed leads on FrontierMath and general math benchmarks; Anthropic's Claude Fable 5 is also competitive, sometimes beating Google, particularly on hardest tiers (claude_news synthesis).
3. **What do related prediction markets imply?** — No closely related Kalshi or Polymarket markets specifically about AI lab rankings were found; keyword searches returned unrelated markets (elections, commodities, geopolitics), so no cross-market corroboration available (kalshi_related, polymarket_related).
4. **Historical base rate?** — Naive uniform base rate across ~7 labs ≈14.3%. Skill-adjusted Monte Carlo (subjective strength scores, OpenAI>Google>>rest) estimates P(Google exactly #2) ≈33-45% under realistic volatility assumptions, with P(Google top-2) ≈76% and P(top-3) ≈91% (code_execution).
# Key facts (high-confidence, factual)
1. [gizmodo/intuitionlabs] Google DeepMind and OpenAI tied at IMO 2025 gold-level (35/42); Google's was officially certified.
2. [kearai.com] Google's Gemini 3 Pro led KEAR Math Arena as of Feb 2026, displacing OpenAI.
3. [llm-stats/benchlm/pricepertoken] By mid-to-late 2026, OpenAI models (GPT-5.5/5.6/5.2 series) top most general/FrontierMath leaderboards; Google not in top 3 of the Aug 2026 general Mathematics leaderboard snapshot.
4. [digitalapplied.com] Anthropic's Claude Fable 5 is statistically tied with or ahead of OpenAI on FrontierMath v2 hardest tier, positioning Anthropic as a strong #2/#3 contender, competing directly with Google.
5. [polymarket_direct] Market priced Google-2nd at 28.5%, sharply down from ~70% a week prior — a large repricing event, likely triggered by a specific benchmark release/leaderboard update not fully detailed in research.
# Cross-market signals
- Kalshi related: No topical matches found for AI lab rankings; keyword searches returned unrelated markets only.
- Polymarket: Self-referential data only (no sister markets on Google/OpenAI/Anthropic math rank found in the "related" scan).
- Sportsbook implied: N/A.
# Analyst opinions and speculation
- Claude_news synthesis concludes ranking is "benchmark-dependent" and a "close call between Google and Anthropic," with OpenAI likely #1 by most current math-specific benchmarks as of August 2026 — implying Google's realistic path to resolution is contesting #2 vs. Anthropic, not #1.
- Code_execution Monte Carlo (subjective, not live-data-based) suggests Google's strength is close to OpenAI's, making Google nearly as likely to be #1 as #2, rarely falling to #3 — but this model uses hand-assigned scores and explicitly is not a live-data pull.
# Directional lean per outcome
- **Yes (Google #2)**: Supported by Google's strong institutional credentials (certified IMO gold, Feb 2026 arena lead) and analyst framing of Google as a top-2/3 fixture. Opposed by multiple mid/late-2026 benchmark snapshots (FrontierMath, general Math leaderboard) showing OpenAI clearly #1 and Google absent from top 3, plus Anthropic's Claude Fable 5 emerging as a strong #2 rival on hardest-tier math. Recent sharp Polymarket price drop (-41.5% in 7 days) also signals market sentiment moving away from Yes.
- **No (Google not #2)**: Supported by the balance of most recent (Jun-Aug 2026) benchmark evidence pointing to OpenAI #1 and either Anthropic or another lab contesting #2, with Google frequently missing top-3 general math leaderboards. Polymarket's current 28.5% price also leans toward No as consensus.
# Gaps / unknowns
- No live check of the actual resolution source (arena.ai Text Arena Math, Labs-filtered Lab Rank) was performed — all evidence is proxy benchmarks (FrontierMath, KEAR, generic "Mathematics" leaderboards), which may not map directly onto arena.ai's specific methodology.
- Cause of the -41.5% 7-day Polymarket price swing is unclear/unsourced in research — could reflect a specific arena.ai update not captured in claude_news.
- No Kalshi-direct YES price was returned in raw research despite instructions to anchor on it; Polymarket used as substitute anchor.
- Thin trading volume ($17.9k) on Polymarket limits confidence in price as true consensus.
# Calibration anchors
- Polymarket current price (proxy anchor): 28.5% YES, down from ~70% a week ago, up slightly over 30 days.
- Skill-adjusted base rate model: ~33-45% for Google exactly #2 under realistic volatility; naive uniform base rate ~14%.
- Precedent: rapid, repeated leadership swaps between OpenAI/Google/Anthropic across different math benchmarks in 2026 suggest high month-to-month volatility, making point-in-time rank forecasts inherently uncertain.