OpenAI’s recent launch of GPT-6 Astra has driven its 52.8% market-implied odds by claiming the top spot on the Code Arena WebDev leaderboard with an 1,797–1,800 Elo score, narrowly ahead of Anthropic’s Claude Fable 5.1. The crowdsourced arena evaluates agentic web development workflows that test multi-step reasoning, tool use, and iterative frontend coding rather than static benchmarks. This positions the latest proprietary large language models from the two leading labs well ahead of Google, Alibaba, Moonshot, and other entrants, whose scores cluster 100–200 points lower. With roughly five weeks until resolution, traders appear focused on whether either company will ship a meaningful update or whether the current gap holds in ongoing community evaluations.
Experimentelle KI-generierte Zusammenfassung mit Polymarket-Daten. Dies ist keine Handelsberatung und spielt keine Rolle bei der Auflösung dieses Marktes. · AktualisiertOpenAI 53.0%
Anthropic 38%
Google 3.3%
Alibaba 1.6%
$76,125 Vol.
$76,125 Vol.

OpenAI
53%

Anthropic
38%

3%

Alibaba
2%

Moonshot
1%

Z.ai
1%

Meta
<1%

SpaceXAI
<1%

Tencent
<1%

ByteDance
<1%

MiniMax
<1%

Xiaomi
<1%

DeepSeek
<1%

Thinky
<1%

Poolside
<1%

Mistral
<1%
OpenAI 53.0%
Anthropic 38%
Google 3.3%
Alibaba 1.6%
$76,125 Vol.
$76,125 Vol.

OpenAI
53%

Anthropic
38%

3%

Alibaba
2%

Moonshot
1%

Z.ai
1%

Meta
<1%

SpaceXAI
<1%

Tencent
<1%

ByteDance
<1%

MiniMax
<1%

Xiaomi
<1%

DeepSeek
<1%

Thinky
<1%

Poolside
<1%

Mistral
<1%
Results from the "Rank" column under the "Code Arena | WebDev" Leaderboard tab at https://arena.ai/leaderboard/code/webdev (Overall) filtered for "Models" will be used to resolve this market.
Note: Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Code Arena | WebDev Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Markt eröffnet: Aug 12, 2026, 7:56 PM ET
Abwickler
0x69c47De9D...Results from the "Rank" column under the "Code Arena | WebDev" Leaderboard tab at https://arena.ai/leaderboard/code/webdev (Overall) filtered for "Models" will be used to resolve this market.
Note: Models marked “AutoEval” at the applicable check time will not be considered, regardless of whether they display a rank or score.
Models will be ordered primarily by their leaderboard rank at the market’s check time. If two or more models are tied on rank, they will be ordered by which model is listed higher on the leaderboard. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the arena.ai Code Arena | WebDev Leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to "Other".
Abwickler
0x69c47De9D...OpenAI’s recent launch of GPT-6 Astra has driven its 52.8% market-implied odds by claiming the top spot on the Code Arena WebDev leaderboard with an 1,797–1,800 Elo score, narrowly ahead of Anthropic’s Claude Fable 5.1. The crowdsourced arena evaluates agentic web development workflows that test multi-step reasoning, tool use, and iterative frontend coding rather than static benchmarks. This positions the latest proprietary large language models from the two leading labs well ahead of Google, Alibaba, Moonshot, and other entrants, whose scores cluster 100–200 points lower. With roughly five weeks until resolution, traders appear focused on whether either company will ship a meaningful update or whether the current gap holds in ongoing community evaluations.
Experimentelle KI-generierte Zusammenfassung mit Polymarket-Daten. Dies ist keine Handelsberatung und spielt keine Rolle bei der Auflösung dieses Marktes. · Aktualisiert
Vorsicht bei externen Links.
Vorsicht bei externen Links.
Häufig gestellte Fragen