
Next Claude Opus Model: Humanity’s Last Exam Debut?
核心摘要
根据「Next Claude Opus Model: Humanity’s Last Exam Debut?」的最新预测市场数据,交易者已形成强烈共识。
目前,45%+ 以压倒性的 92.5% 获胜概率主导市场;50%+ 以 86.9% 位居第二,55%+ 以 45% 排名第三。该市场的下注量已达 $92.1K,反映出市场的高度关注。
竞争梯队拆解
为了更好地评估各潜在结果的位置,可依据隐含概率与合约定价将市场划分为三个明显的交易梯队:
🥇 第一梯队:绝对领跑者
- 45%+ (92.5%):45%+ 目前拥有最高概率,深受订单簿青睐。看好该结果的交易者面对的「Buy Yes」合约价为 93¢,显示出市场的高度确信。仅该合约就已产生 $27.1K 的成交量。
🥈 第二梯队:主要挑战者
- 50%+ (86.9%):作为最可行的替代选项,50%+ 保持着 86.9% 的成真概率,其「Buy Yes」份额目前成交价为 87¢。
- 55%+ (45%):以 45% 的概率位列第三,市场对 55%+ 持谨慎怀疑态度,除非势头转变,否则视其为外围黑马。
🥉 第三梯队:长尾选项(合计约 0%)
在前三名之外,还有大量宏观变量与冷门结果被持续追踪。尽管单个概率偏低,但它们是投机交易者的重要对冲:
- 替代选项:包括 60%+ (12%)。
- 投机成交:尽管统计概率偏低,像 60%+ 这类长尾合约仍吸引着可观的关注。
完整订单簿与定价面板
下表列出了该预测池中所有结果的合约价格、概率与市场深度的完整拆解:
| 排名 | 预测结果 | 获胜概率 | 成交量 | 买入 Yes(成本) | 买入 No(成本) |
|---|---|---|---|---|---|
| 1 | 45%+ | 92.5% | $27.1K | 93¢ | 8¢ |
| 2 | 50%+ | 86.9% | $5.9K | 87¢ | 13¢ |
| 3 | 55%+ | 45.0% | $20.5K | 45¢ | 55¢ |
| 4 | 60%+ | 12.0% | $38.6K | 12¢ | 88¢ |
裁决规则
This market will resolve to "Yes" if the next Anthropic Claude Opus model added to the Humanity’s Last Exam results at https://agi.safe.ai/ has an HLE Accuracy of at least the specified percentage at 12:00 PM ET on the calendar date following the date on which it first appears on the site. Otherwise, this market will resolve to "No".
Any Anthropic Claude model newly added to the Humanity’s Last Exam results and labeled as "Opus" may qualify (e.g., claude-opus-5.1, claude-opus-5.5-thinking, claude-opus-6-preview, or similar). Claude models labeled only as Sonnet, Haiku, Fable, Mythos, or another non-Opus variant will not qualify.
The percentage displayed as "HLE Accuracy" for the model in its result card on the "AI Progress on Humanity’s Last Exam" chart at https://agi.safe.ai/ will be used to resolve this market. This market will resolve solely based on the displayed HLE Accuracy, regardless of the model’s Calibration Error, chart position, other configuration scores, or any underlying granular or unrounded data presented elsewhere.
If multiple qualifying models are added to the Humanity’s Last Exam results on the same calendar date (ET), the model with the highest HLE Accuracy will be used for resolution. Models added to the results on the calendar date following the initial qualifying model’s first appearance will not be considered. If a model is first added to the Humanity’s Last Exam results but is subsequently removed and not re-added, such that it is not displayed on the site at 12:00 PM ET on the calendar date following its first appearance, its appearance will not qualify as added to the Humanity’s Last Exam results.
A qualifying model must be newly added to the Humanity’s Last Exam results at https://agi.safe.ai/. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the site is irrelevant for this market.
The resolution source for this market is the official Humanity’s Last Exam website found at https://agi.safe.ai/. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the site, this market will resolve based on the first subsequent instance at which the model’s HLE Accuracy becomes available on the site. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the site or if no qualifying model is added by December 31, 2026, 11:59 PM ET, this market will resolve to "No".
AI 估值分析:发现市场错误定价与 EV 差
人群共识与投机成交塑造了更宏观的预测市场,而我们的量化算法提供了数据驱动的反向视角。通过分析基本面信号、底层趋势与历史分布,我们的 AI 估值模型为每个结果独立测算出一个「公允价值」概率。
将该公允价值与当前交易价值对比,可揭示出重大背离——即期望值(EV)差。正 EV 差代表统计上被低估的结果,而负 EV 差则提示市场可能存在反应过度。
顶级 AI Alpha 与错误定价套利机会
根据最新一轮数据模型测算,以下几个关键合约存在显著偏离:
- 最被高估的结果:60%+ 当前交易价为 12%,但我们的 AI 测算其公允价值仅为 10.7%,形成 -1.3% 的较大负 EV 差,表明人群可能过度炒作该结果、把溢价推得过高。
- 最佳价值标的(最高 EV):我们的模型将 50%+ 识别为盘面上最具价值的机会。市场仅给予其 86.9% 的交易概率,而我们 AI 的公允价值评估为 90.2%——形成可观的 +3.3% EV 差。
- 被忽视的黑马:其他值得注意的偏离包括 55%+(EV 差:+0.8%)。尽管我们的预测模型给予更强的统计支撑,这些长尾机会仍被实时订单簿大幅低估。
| Market | Trade Value | Fair Value | EV Gap |
|---|---|---|---|
| 45%+ | 92.5% | 93.6% | +1.1% |
| 50%+Best EV | 86.9% | 90.2% | +3.3% |
| 55%+ | 45.0% | 45.8% | +0.8% |
| 60%+ | 12.0% | 10.7% | -1.3% |
交易动态
以下是该事件的交易动态。
Sep 26, 2026
- 08:23 PMAJAJSV$1.92
Bought 5.18919 Yes for Will the next Claude Opus model debut with a Humanity’s Last Exam score of 55% or higher? at 0.37
- 08:20 PMEIeightpenguins$282.11
Bought 300.12 No for Will the next Claude Opus model debut with a Humanity’s Last Exam score of 60% or higher? at 0.94
- 08:20 PMFEFelicityBright$3.52
Sold 44 No for Will the next Claude Opus model debut with a Humanity’s Last Exam score of 45% or higher? at 0.08
- 08:19 PMPRprjjfb34$3.67
Sold 40.83 No for Will the next Claude Opus model debut with a Humanity’s Last Exam score of 45% or higher? at 0.09
- 08:19 PMNIninasimon$1.34
Bought 1.467689 Yes for Will the next Claude Opus model debut with a Humanity’s Last Exam score of 45% or higher? at 0.91
- 08:19 PMTRtrader-04ad8cb0$4.55
Bought 5 Yes for Will the next Claude Opus model debut with a Humanity’s Last Exam score of 45% or higher? at 0.91
- 08:19 PMGOGollumGekko$222.52
Bought 244.53 Yes for Will the next Claude Opus model debut with a Humanity’s Last Exam score of 45% or higher? at 0.91
- 08:18 PMPIpieped$90.48
Bought 96.25 No for Will the next Claude Opus model debut with a Humanity’s Last Exam score of 60% or higher? at 0.94
- 08:17 PMPIpieped$1,373.14
Bought 1460.79 No for Will the next Claude Opus model debut with a Humanity’s Last Exam score of 60% or higher? at 0.94
- 08:17 PMPIpieped$148.68
Bought 252 No for Will the next Claude Opus model debut with a Humanity’s Last Exam score of 55% or higher? at 0.59
- 08:16 PMPIpieped$62.40
Bought 120 No for Will the next Claude Opus model debut with a Humanity’s Last Exam score of 55% or higher? at 0.52
- 08:15 PMSHshoobadooba18$6.16
Sold 88 Yes for Will the next Claude Opus model debut with a Humanity’s Last Exam score of 60% or higher? at 0.07
正在押注该事件的鲸鱼钱包
常见问题
当前市场对「Next Claude Opus Model: Humanity’s Last Exam Debut?」的共识是什么?
截至最新更新,45%+ 以 92.5% 的获胜概率领跑,其次是 50%+(86.9%),以及 55%+(45%)。该市场总成交量已达 $92.1K,显示出充足的流动性与高交易参与度。
AI 公允价值与实时市场交易价值有何不同?
实时市场交易价值反映的是公众情绪、订单簿动能与投机资金。我们的 AI 公允价值则由量化模型独立计算,剔除情绪炒作、专注底层数据。两者出现显著背离时即形成 EV 差,提示市场对某个结果可能存在错误定价。
当前哪个结果的期望值(EV)最高?
最新一轮测算显示,50%+ 是最显著的错误定价。市场对其隐含概率仅给到 86.9%,而我们的 AI 测算其公允价值为 90.2%——形成 +3.3% 的期望值差,是该市场中最具价值的标的。
市场共识是否对某个结果反应过度?
是的——数据显示市场对 60%+ 存在明显的反应过度。人群把其实时交易价值推高至 12%,但我们的公允价值评估认为其真实概率仅为 10.7%,形成 -1.3% 的负 EV 差,表明该合约被高估。
长尾数据中是否藏有高价值的黑马选项?
当然有。除了头部结果之外,我们的模型在排名靠后的选项中发现了被低估的潜力。55%+ 拥有 +0.8% 的正 EV 差。尽管量化层面更有支撑,这些合约仍被实时订单簿低估。
