Anthropic’s September 2026 launch of Claude Fable 5.1 and Mythos 5.1 has driven the latest gains on Humanity’s Last Exam, with Fable 5.1 posting the top reported scores of 59.1% on standard leaderboards and up to 65% in tool-augmented evaluations. These iterative releases emphasize enhanced reasoning, adaptive thinking, and tool integration on the 2,500-question benchmark spanning graduate-level math, science, and humanities. Claude variants now lead or cluster at the frontier ahead of OpenAI’s GPT-6 Astra and other competitors, reflecting sustained investment in frontier capabilities. Traders are watching for additional model updates or optimization techniques before year-end that could push the highest Claude score higher while the benchmark remains far from saturation.
Экспериментальная сводка, созданная ИИ на основе данных Polymarket. Это не является торговой рекомендацией и не влияет на то, как разрешается этот рынок. · ОбновленоСамый высокий балл Клода на последнем экзамене человечества в 2026 году?
$96,744 Объем
55%+
88%
60%+
51%
65%+
21%
70%+
7%
75%+
5%
$96,744 Объем
55%+
88%
60%+
51%
65%+
21%
70%+
7%
75%+
5%
For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Открытие рынка: Jul 23, 2026, 6:42 PM ET
Источник определения исхода
https://agi.safe.ai/Кто определяет исход
0x65070BE91...For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Источник определения исхода
https://agi.safe.ai/Кто определяет исход
0x65070BE91...Anthropic’s September 2026 launch of Claude Fable 5.1 and Mythos 5.1 has driven the latest gains on Humanity’s Last Exam, with Fable 5.1 posting the top reported scores of 59.1% on standard leaderboards and up to 65% in tool-augmented evaluations. These iterative releases emphasize enhanced reasoning, adaptive thinking, and tool integration on the 2,500-question benchmark spanning graduate-level math, science, and humanities. Claude variants now lead or cluster at the frontier ahead of OpenAI’s GPT-6 Astra and other competitors, reflecting sustained investment in frontier capabilities. Traders are watching for additional model updates or optimization techniques before year-end that could push the highest Claude score higher while the benchmark remains far from saturation.
Экспериментальная сводка, созданная ИИ на основе данных Polymarket. Это не является торговой рекомендацией и не влияет на то, как разрешается этот рынок. · Обновлено



Не доверяй внешним ссылкам.
Не доверяй внешним ссылкам.
Часто задаваемые вопросы