Recent releases have intensified competition in coding arenas like Code Arena WebDev and LMArena Coding, where GPT-6 Astra from OpenAI recently claimed the top spot at 1797 points, edging Anthropic's Claude Fable 5.1 (1762) and Opus 5 variants. Anthropic models continue to lead on SWE-bench Verified (up to 96%) and agentic benchmarks such as Terminal-Bench, reflecting strong performance on real GitHub issues and multi-step workflows. With roughly 3.5 months until year-end, trader sentiment centers on whether iterative updates, new scaffolding, or fine-tunes from OpenAI, Anthropic, or Alibaba's Qwen series can push any model past the implied threshold amid rapid capability gains and benchmark saturation. Key catalysts include potential fall model drops and developer conference announcements that could shift Elo ratings or pass rates quickly.
สรุปจาก AI ทดลองที่อ้างอิงข้อมูลจาก Polymarket ไม่ใช่คำแนะนำในการเทรดและไม่มีผลต่อการตัดสินตลาดนี้ · อัปเดตแล้ว$209,644 ปริมาณ
1560
40%
1580
25%
1600
13%
$209,644 ปริมาณ
1560
40%
1580
25%
1600
13%
Results from the "Score" column under the "Text Arena | Coding" Leaderboard tab at https://arena.ai/leaderboard/text/coding-no-style-control with style control off will be used to resolve this market.
The resolution source for this market is the Chatbot Arena LLM Leaderboard found at arena.ai/leaderboard/text. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If permanently unavailable, this market will resolve to "No".
ตลาดเปิดเมื่อ: Apr 2, 2026, 6:09 PM ET
ผู้ตัดสินผล
0x65070BE91...Results from the "Score" column under the "Text Arena | Coding" Leaderboard tab at https://arena.ai/leaderboard/text/coding-no-style-control with style control off will be used to resolve this market.
The resolution source for this market is the Chatbot Arena LLM Leaderboard found at arena.ai/leaderboard/text. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If permanently unavailable, this market will resolve to "No".
ผู้ตัดสินผล
0x65070BE91...Recent releases have intensified competition in coding arenas like Code Arena WebDev and LMArena Coding, where GPT-6 Astra from OpenAI recently claimed the top spot at 1797 points, edging Anthropic's Claude Fable 5.1 (1762) and Opus 5 variants. Anthropic models continue to lead on SWE-bench Verified (up to 96%) and agentic benchmarks such as Terminal-Bench, reflecting strong performance on real GitHub issues and multi-step workflows. With roughly 3.5 months until year-end, trader sentiment centers on whether iterative updates, new scaffolding, or fine-tunes from OpenAI, Anthropic, or Alibaba's Qwen series can push any model past the implied threshold amid rapid capability gains and benchmark saturation. Key catalysts include potential fall model drops and developer conference announcements that could shift Elo ratings or pass rates quickly.
สรุปจาก AI ทดลองที่อ้างอิงข้อมูลจาก Polymarket ไม่ใช่คำแนะนำในการเทรดและไม่มีผลต่อการตัดสินตลาดนี้ · อัปเดตแล้ว



ระวังลิงก์ภายนอก
ระวังลิงก์ภายนอก
คำถามที่พบบ่อย