Meta’s Muse Spark 1.1 series, with scores reaching 52–62% on September 2026 Humanity’s Last Exam leaderboards depending on tool use and reasoning modes, forms the core driver of trader sentiment. These results place Meta models ahead of several GPT-5 variants and competitive with early Gemini releases while trailing Anthropic’s leading Claude Fable and Opus entries at 59–65%. Rapid frontier progress on the expert-authored 2,500-question benchmark—spanning graduate-level math, physics, and multimodal tasks—reflects ongoing architectural gains in reasoning and post-training. Key catalysts through year-end include potential Muse Spark refinements or successor releases that could lift Meta’s highest 2026 score toward or past 60%, alongside any shifts in evaluation protocols or private test-set performance that affect official resolution.
Résumé expérimental généré par IA à partir des données Polymarket. Ceci n'est pas un conseil de trading et ne joue aucun rôle dans la résolution de ce marché. · Mis à jour$46,098 Vol.
55 %+
44%
60 % ou plus
28%
65 %+
18%
70 %+
7%
$46,098 Vol.
55 %+
44%
60 % ou plus
28%
65 %+
18%
70 %+
7%
For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Marché ouvert : Jul 23, 2026, 6:48 PM ET
Source de résolution
https://agi.safe.ai/Résolveur
0x65070BE91...For resolution, “accuracy” refers solely to the value labeled “HLE Accuracy” or a clear equivalent metric if the data’s presentation or terminology is restructured, regardless of the model’s Calibration Error or any other displayed metric.
The resolution source will be the official Humanity’s Last Exam leaderboard at https://agi.safe.ai/. If the resolution source becomes unavailable during the listed timeframe, this market will remain open to allow the relevant data to become available again. If the source remains unavailable after the end of the listed timeframe or is otherwise confirmed to be permanently unavailable, official Humanity’s Last Exam results published elsewhere may be used. If no official alternative source is available, this market will resolve to "No".
Source de résolution
https://agi.safe.ai/Résolveur
0x65070BE91...Meta’s Muse Spark 1.1 series, with scores reaching 52–62% on September 2026 Humanity’s Last Exam leaderboards depending on tool use and reasoning modes, forms the core driver of trader sentiment. These results place Meta models ahead of several GPT-5 variants and competitive with early Gemini releases while trailing Anthropic’s leading Claude Fable and Opus entries at 59–65%. Rapid frontier progress on the expert-authored 2,500-question benchmark—spanning graduate-level math, physics, and multimodal tasks—reflects ongoing architectural gains in reasoning and post-training. Key catalysts through year-end include potential Muse Spark refinements or successor releases that could lift Meta’s highest 2026 score toward or past 60%, alongside any shifts in evaluation protocols or private test-set performance that affect official resolution.
Résumé expérimental généré par IA à partir des données Polymarket. Ceci n'est pas un conseil de trading et ne joue aucun rôle dans la résolution de ce marché. · Mis à jour



Méfiez-vous des liens externes.
Méfiez-vous des liens externes.
Questions fréquentes