跳至正文
目录

母裁决:AGI 的预言史是真的、专家调查的数字是真的、预测市场的定价是真的、行业高管的时间表是真的——而这一切的真实都挂在一个从未被权威地画下来的名词上。1950 年图灵把问题换成游戏、1997 年 Gubrud 造出全称、2018 年 OpenAI 章程给出「经济上胜过人类」、2023 年 DeepMind 划出五级、2024 年微软-OpenAI 协议把它定价为一千亿美元利润、2025 年 Hendrycks 又把门槛搬到「认知上与受过良好教育的成年人相当」——同一个词,六把尺,刻度之间相差几十年。历史记分卡显示:经典预言全部未兑现,近年「2–5 年」窗口是历史上第一次从业者与研究者同时押注短窗口;专家调查中位数在一年内前移 13 年而专家自认分歧很大;预测市场把 50% 概率从「五十年后」压到「十年内」又往外拉回;而监管文件与上市公司年报里,这个词出现的次数是零。「AGI 何时到来」未立,「AGI 是纯炒作、永远不会来」同样未立——真正可判定的是第三件事:AGI 从来不是一个有待兑现的日期,而是一个等待被定义的承诺;问「还差几年」之前,先问「你说的 AGI 是哪一个」。

纪律红线:本篇不预测 AGI 年份,不构成任何投资、就业或政策建议;不评价任何在世研究者的品格;把「定义不清」裁成测量问题而非道德问题;把预测市场的时点价与专家调查的百分位当作「带口径的条件语句」而不是「科学概率」。

零、先看结论

0.1 九项裁决

命题 裁决 置信度
「AGI」是一个有明确出生证、可考古的词 成立。 Turing 1950 提出模仿游戏;Gubrud 1997 首次使用全称;Goertzel 圈子 2002 定书名、2008 办首次会议
AGI 有一个被普遍接受的权威定义 不成立。 经济口径(OpenAI 章程)、能力口径(DeepMind 五级)、合同口径($100B 利润)、认知口径(Hendrycks 2025)并行,刻度差数十年;联邦文件与年报中出现 0 次
1950s–1970s 的经典 AGI 预言今天兑现了 不成立。 Simon 1957 四条十年预言全落空(象棋拖了约 40 年);Minsky 1967「一代人内解决」未兑现;Minsky 1970「3–8 年」被本人 1996 年否认;Lighthill 1973 判词短期生效
2023–2026 行业高管的「2–5 年」窗口有统一口径 不成立。 Altman「whooshing by」、Amodei「powerful AI 最早 2026」、Hassabis「2030 前后」、Musk 年年顺延——术语与口径互不通用
专家调查的「中位年份」是稳定的科学结论 不成立。 2022 年中位 2059(后修订 2060)→2023 年 2047,一年前移 13 年;同卷「年份题 vs 概率题」答案差一倍;专家自认与他人分歧「not much」占 44%
预测市场的 AGI 定价是科学概率 不成立。 Metaculus 主问题用四条硬指标定义(2 小时对抗图灵测试/组装法拉利模型/90% 准确率/APPS 90.0%),不是「AGI」通义;2020 中位「50 年后」→2026-02「25% by 2029/50% by 2033」,口径不变而定价随事件游走
AGI 作为合同触发词存在 曾经成立、已死亡。 微软-OpenAI 协议把 AGI 定义为「能产生约 $100B 利润的系统」;2026-04-27 协议重签,AGI 条款被固定日期(2030/2032)取代
2024–2026 已有人/公司宣布 AGI 并获确认 不成立。 2025 年主要实验室宣称「2025 AGI」类预测在 Metaculus 定价 0%、Polymarket 全部解决 No;ARC-AGI-2 的 85% 门槛无人达到(榜首 24.03%)
因为定义不清、预测常错,所以 AGI 是纯炒作、永远不来 不成立。 能力轨迹真实可测(ARC 从 0%→54%、HLE 从 <10%→64.7%、IMO 金牌);虚无派自己的路线之争(LeCun 反 LLM)同样是无证主张

0.2 一句话地图

层级 截至 2026-08-01 能说到哪里 不能再多走的一步
词源 AGI 一词有明确出生证(Gubrud 1997→Goertzel 2002/2008),但从未有统一的实体指称 不能把「词的存在」读成「物的存在」
定义 至少六套并行口径,刻度从 2027 走到 2116 不能挑一把尺回答「何时到来」
预言 经典预言全落空;2023–2026 窗口是「从业者+研究者」同押短窗口的首例 不能把高管访谈当官方承诺,也不能把落空当「必然再落空」
调查 2023 轮中位数 2047/10% by 2027;专家分歧大、框架效应差一倍 不能把聚合中位数当科学概率
市场 Metaculus 2026-02「25% by 2029/50% by 2033」;2025 年 AGI 宣称类市场全部解决 No 不能把时点价当事件概率
合同 AGI 曾是合同触发词($100B 利润),2026-04-27 废止 不能把合同的死说成概念的死
监管 EO 14179、NIST 词汇表、TTC 术语表对 AGI 零出现 不能把「官方没定义」说成「官方否认」
能力 ARC-AGI-2 从 0%→54%(refinement),HLE 从 <10%→64.7%,轨迹真实可测 不能把「在逼近」读成「已到达」
虚无侧 LeCun/Marcus 的批评有真论点,其路线判断本身是无证主张 不能把批评者的存在读成「AGI 必不来」

0.3 证据标签

  • 一手逐字: 本次抓取或经权威转录核对的原文;
  • 文献较稳: 多个独立一手来源一致的结论;
  • 理论整合/推断: 本篇把多条证据连接后的解释;
  • 我们的断言: 明示所用门槛的裁决,不冒充机构原话;
  • 需亲核: 二手转述、转录级或动态页面,已如实标注。

一、完善后的课题与事前承诺

1.1 三件可判定的事

本篇不问「AGI 哪年来」(不可判定,且每一把尺给出不同答案),只问三件可查证的事:

  1. AGI 在写下它的文件里是什么? 定义谱系:Turing 1950 → Gubrud 1997 → Goertzel 2002/2008 → OpenAI 章程 2018 → DeepMind 五级 2023 → $100B 协议 2024 → Hendrycks 2025。
  2. 写下过年份的预言今天对不对得上? 历史记分卡(1956–2026 逐条预言、原文与兑现),加上 2023–2026 公司高管时间表的谱系。
  3. 预测市场的 AGI 概率是不是概率? Metaculus/Manifold/Polymarket 的定价、解决标准与历史游走,专家调查中位数与其方法论批评。

1.2 四次升格与一次反向抹除

  • 日期跳: 把「能力在推进」写成「AGI 在 X 年到来」——每年给出一个具体年份,却从不写该年份对应的完工判据;
  • 定义跳: 把「某个具体模型的能力」写成「AGI 已实现」——2024–2026 的 o1/GPT-5/Claude 宣称与反驳之战;
  • 概率跳: 把预测市场均衡价/专家中位数当成「科学概率」——带口径的条件语句被读成对世界的断言;
  • 名号跳: 同一个词在能力、经济、合同、认知四口径间自由滑动,两侧都用它赢下论点;
  • 反向抹除: 因为定义不清、预测常错、宣称常落空,便断言「AGI 是宗教、永远不会来」——把测量困难反向跳成本体否定。

1.3 预先写下的分界声明

  • 与技术奇点篇(2026-07-03)分界: 那篇审「趋势→命定/能力→无界标量」的概念本体;本篇不重审奇点概念,只审 AGI 年份预言、概率定价与定义谱系的记账与兑现。奇点篇已用的 Minsky 1970/Kurzweil 2045 引文,本篇引用时注明出处并做记分卡处理,不复述其论证。
  • 与 Scaling 撞墙篇(2026-07-23)分界: 那篇审 2024Q4–2026 单轴放缓论战与标度律本身;本篇不重审能力轨迹,只把 ARC/HLE/IMO 作为「可测能力刻度」用作判据账的锚。
  • 与预测市场篇(2026-07-20)分界: 那篇已裁「价格≠真概率」;本篇直接复用它,把预测市场当作 AGI 预言的数据点之一,不再重做价格本体审计。
  • 与 AI 就业篇(2026-07-21)分界: 不重做 TFP/替代账。
  • 与 LLM 推理/ICL/Agent 篇分界: 不重审「模型会不会推理」,只审「推理能力的进展被读成了什么预言」。
  • 与 L5/固态电池/抗衰路线图/核聚变篇(技术时间表家族)的关系: 本篇是该家族在 AI 预言侧的应用,但有一处根本差异——那几篇的病是「判据在场而口径缺席」或「没人朝终点报进度」,而 AGI 连一份权威完工判据都从未存在过:判据账比日期账更根本。

1.4 预先写下的改判条件

  • 若出现一份被多方承认的 AGI 完工判据(如监管或行业共识定义),且该判据下的年份预言持续兑现,升级「定义」裁决;
  • 若 ARC-AGI-2/3 的 85% 门槛被开源、可复现系统达成,升级「能力刻度」判断;
  • 若专家调查显示中位数多年稳定且框架效应消失,升级「调查」判断;
  • 若预测市场在固定口径下长期校准良好,升级「市场」判断;
  • 若 2026–2027 高管窗口落空,记分卡追加一条「从业者窗口失效」记录——但单次落空不构成「永不来」证据。

1.5 明确不做

本篇不预测 AGI 年份;不把任何公司高管的访谈原话当作正式承诺;不把「定义不清」写成对任何人的道德指控;不重写奇点论证、scaling 论战或 AI 安全争议;不以个别落空的预言否定 AI 的真实进展;不以「预测常错」为由否定预测本身的价值(那正是本库预测市场篇裁过的反向越界)。

二、定义谱系考古:一个词的出生证

2.1 Turing 1950:问题被换成游戏

Alan Turing 1950 年在《Mind》发表的《Computing Machinery and Intelligence》没有使用 “artificial intelligence” 或 “AGI”,使用的是 “thinking machines” / “machine intelligence”,并且明确拒绝原问题。[一手逐字] 开头原句:「I propose to consider the question, ‘Can machines think?’ This should begin with definitions of the meaning of the terms ‘machine’ and ‘think.’」问题随即被替换为模仿游戏:「We now ask the question, ‘What will happen when a machine takes the part of A in this game?’…These questions replace our original, ‘Can machines think?’」[一手逐字] 著名预言(§6):「I believe that in about fifty years’ time it will be possible, to programme computers, with a storage capacity of about 10^9, to make them play the imitation game so well that an average interrogator will not have more than 70 per cent chance of making the right identification after five minutes of questioning.」[一手逐字] 来源:Turing 1950 全文抄录本(umbc)牛津 Mind 出版记录

理论整合/推断: 图灵已经把「能不能想」换成「一个受控游戏中能否骗过裁判」——可判定的第一步,也正是后来一切「AGI 定义之争」的原型:每一条判定标准都在定义里预先埋下了被挑战的接口。

2.2 Gubrud 1997:全称的首次出现

Mark Gubrud 1997 年 11 月在第五届 Foresight 分子纳米技术会议提交的论文《Nanotechnology and International Security》中首次使用 “artificial general intelligence”。[一手逐字] 摘要首句:「However, assembler-based nanotechnology and artificial general intelligence have implications far beyond the Pentagon’s current vision of a ‘revolution in military affairs.’」[一手逐字] 定义段:「By advanced artificial general intelligence, I mean AI systems that rival or surpass the human brain in complexity and speed, that can acquire, manipulate and reason with general knowledge, and that are usable in essentially any phase of industrial or military operations where a human intelligence would otherwise be needed.」[一手逐字] 来源:Wayback 快照(原文页已下线)

我们的断言(考古勘误一): Gubrud 是描述性使用,未自称「造词者」;「AGI」缩写的普及来自 Goertzel 圈子(书名与会议),而通行说法「Goertzel 2002 年办首届 AGI 会议」不实——2002 年只是为 2007 年 Springer 文集定名(书名由 Shane Legg 提议),首个 AGI 研讨会是 2006 年(Bethesda),首个正式 AGI 会议是 2008 年(孟菲斯大学)。[一手逐字] Goertzel 2013 年 MIRI 访谈:「We put together the first AGI Workshop in Bethesda in 2006…the first full-on AGI conference was in 2008 at the University of Memphis」;[一手逐字] 官方会议报告开篇:「The First Conference on Artificial General Intelligence (AGI-08) was held on March 1–3, 2008, at the University of Memphis.」来源:MIRI 访谈AI Magazine 会议报告

2.3 OpenAI 章程 2018:经济口径

考古勘误二: 「economically valuable work」定义不出自 2015–2016 年创立文件。2015-12-11《Introducing OpenAI》原文没有 AGI 定义,只用 “digital intelligence” / “human-level AI”:「Our goal is to advance digital intelligence in the way that is most likely to benefit humanity as a whole, unconstrained by a need to generate financial return.」[一手逐字] 来源:Introducing OpenAI(Wayback)

现行定义出自 2018 年 4 月首次发布的章程(主笔 2026-08-01 亲核现行页):[一手逐字]「OpenAI’s mission is to ensure that artificial general intelligence (AGI)—by which we mean highly autonomous systems that outperform humans at most economically valuable work—benefits all of humanity.」[一手逐字] 章程自带附带条款:「if a value-aligned, safety-conscious project comes close to building AGI before we do, we commit to stop competing with and start assisting this project…a typical triggering condition might be ‘a better-than-even chance of success in the next two years.’」来源:OpenAI Charter

理论整合/推断: 这是「经济口径」的正源:AGI 被定义为经济上胜过人类——一个没有能力清单、没有可测门槛、只有相对标准的定义。它后来在微软-OpenAI 合同里被进一步实体化为利润数字(见第五节)。

2.4 Chollet 2019:能力谱系口径

François Chollet《On the Measure of Intelligence》(arXiv:1911.01547)给出与「AGI 是二元事件」直接冲突的定义:[一手逐字] 摘要:「We then articulate a new formal definition of intelligence based on Algorithmic Information Theory, describing intelligence as skill-acquisition efficiency and highlighting the concepts of scope, generalization difficulty, priors, and experience.」[一手逐字] 正文 §II.2.1:「The intelligence of a system is a measure of its skill-acquisition efficiency over a scope of tasks, with respect to priors, experience, and generalization difficulty.」[一手逐字]「general intelligence」不是二元属性:「a central point of this document is that ‘general intelligence’ is not a binary property which a system either possesses or lacks. It is a spectrum…」来源:arXivHTML 版

需亲核: 网上流传的「Intelligence is not a skill, but rather the efficiency with which an agent acquires new skills」不在论文正文(主笔对 ar5iv 全文本检索未命中),更可能是访谈转述——引用时只用论文原文。

ARC-AGI 的官方定位(同论文 §6):[一手逐字]「ARC can be seen as a general artificial intelligence benchmark, as a program synthesis benchmark, or as a psychometric intelligence test.」来源:arXiv

2.5 DeepMind 五级 2023:能力-通用性-自主三维口径

Morris、Sohl-dickstein、Legg 等《Levels of AGI》(arXiv:2311.02462)把 AGI 从「有没有」改成「第几级」:[一手逐字] 摘要:「This framework introduces levels of AGI performance, generality, and autonomy, providing a common language to compare models, assess risks, and measure progress along the path to AGI.」五级逐字(v5 版):Level 1「Emerging: equal to or somewhat better than an unskilled human」;Level 2「Competent: at least 50th percentile of skilled adults」;Level 3「Expert: at least 90th percentile of skilled adults」;Level 4「Exceptional: at least 99th percentile of skilled adults」;Level 5「Superhuman: outperforms 100% of humans」。来源:arXiv

我们的断言(考古勘误三): OpenAI 的「五级路线图」(Chatbots→Reasoners→Agents→Innovators→Organizations)是 2024 年 7 月 Bloomberg 报道的,不是 2023 年;且 Altman 同年 11 月已表态弃用二元「是/否 AGI」:「We try now to use these different levels…rather than the binary of, ‘is it AGI or is it not?’ I think that became too coarse as we get closer.」[多源交叉] 来源:Bloomberg 2024-07-11存档)、Axios 复述

2.6 HLE 与 ARC Prize 2025:有门槛的尺

  • Humanity’s Last Exam(CAIS + Scale AI,2024-09 发布):[一手逐字]「designed to be the final closed-ended academic benchmark of its kind…The dataset consists of 2,500 challenging questions across over a hundred subjects.」没有官方及格线;论文自陈即使模型 2025 年底超 50%:「it would not alone suggest autonomous research capabilities or ‘artificial general intelligence.’」来源:HLE 公告Nature 论文 s41586-025-09962-4
  • ARC Prize 2025(ARC-AGI-2):85% 是有官方门槛的尺:[一手逐字]「Your objective: Reach 85% accuracy on the ARC-AGI-2 private evaluation dataset within the Kaggle efficiency limits.」[一手逐字] 发布说明:「Pure LLMs score 0% on ARC-AGI-2, and public AI reasoning systems achieve only single-digit percentage scores. In contrast, every task in ARC-AGI-2 has been solved by at least 2 humans in under 2 attempts.」结果(2025-12-05 官方公告,主笔亲核):Grand Prize 无人领取;[一手逐字]「1,455 teams submitted 15,154 entries…The top Kaggle score winner reached a new SOTA on the ARC-AGI-2 private dataset of 24% for $0.20/task」;商业模型:「the top verified commercial model, Opus 4.5 (Thinking, 64k), scores 37.6% for $2.20/task. The top verified refinement solution, built on Gemini 3 Pro and authored by Poetiq, scores 54% for $30/task.」来源:ARC Prize 2025 条款ARC-AGI-2 发布说明结果分析(Mike Knoop, 2025-12-05)

理论整合/推断: HLE 没有及格线、ARC 有及格线但 85% 无人达到——「AGI 门槛」的现实是:造尺人自己也没法把尺的刻度对齐。ARC Prize 官方在 2025 年结果分析里自陈:ARC-AGI-1/2 的私有测试可能已被「训练到测试」式污染(Gemini 3 验证中的「correct ARC color mappings」证据),且「We still need new ideas, like how to separate knowledge and reasoning」。造尺人自己说「So, do we have AGI? Not yet.」[一手逐字]

2.7 Hendrycks 2025:认知口径与「移动靶」的自认

Hendrycks、Marcus、Bengio、Schmidt 等 2025-10《A Definition of AGI》(arXiv:2510.18212):[一手逐字]「Artificial General Intelligence (AGI) may become the most significant technological development in human history, yet the term itself remains frustratingly nebulous, acting as a constantly moving goalpost.」其操作化定义:[一手逐字]「AGI is an AI that can match or exceed the cognitive versatility and proficiency of a well-educated adult.」并给出量化分数(GPT-4 27%、GPT-5 57%)。来源:arXiv

「移动靶」不是批评者的发明——Altman 自己在访谈里承认([多源交叉] 经 LessWrong 转录转引,原访谈出处未核):「Now we’re going to move the goalposts, always, which is why this is hard, but I’ll stick with that as an answer.」;OpenAI 官方 2025 年博客:[多源交叉]「We used to view the development of AGI as a discontinuous moment…We now view the first AGI as just one point along a series of systems of increasing usefulness.」来源:LessWrong 转录the-decoder 转引

三、历史预言记分卡:1956–2026

3.1 记分卡总表

预言者/年份 原文核心句(逐字) 兑现
Newell & Simon 1957/1958 「within ten years a digital computer will be the world’s chess champion, unless the rules bar it from competition」等四条 全部未兑现;象棋拖到 1997(约 40 年),「发现并证明重要新数学定理」至今无公认案例
Simon 1960/1965 「machines will be capable, within twenty years, of doing any work that a man can do」 1980/1985 均未兑现
Minsky 1967 「Within a generation, I am convinced, few compartments of intellect will remain outside the machine’s realm—the problems of creating ‘artificial intelligence’ will be substantially solved」 一代人(约 1995)内未兑现
Minsky 1970(Life) 「In from three to eight years we will have a machine with the general intelligence of an average human being」 未兑现;本人 1996 年否认数字(见 3.2)
Lighthill 1973 「Most workers in AI research and in related fields confess to a pronounced feeling of disappointment in what has been achieved in the past twenty-five years」 短期生效(引爆第一次 AI 冬天);「通用机器人注定失败」长期被翻案
Moravec 1988/1993-94 「Within forty years…we will achieve human equivalence in our machines」;「By around 2040, there will be no job that people can do better than robots」 2028/2040 均未兑现
Kurzweil 1999/2005 「By 2029 the software for intelligence will have been largely mastered」;「I set the date for the Singularity…as 2045」 2029 未到、2045 未到;Kurzweil 2024 自评 2029 预测「可能保守」(二手转引)
Musk 2014–2026 「five year time frame, ten years at most」(2014)→「probably next year, within two years」(2024)→「AGI next year in ’26」(2026-01) 唯一「年年顺延、已有确定落空记录」的序列:2024 年「AGI by 2025」已落空
Altman 2023–2026 「AGI could happen soon or far in the future」(2023-02 官方博客)→「confident we know how to build AGI」(2025-01)→「AGI kinda went whooshing by」(2025-12,转录级) 从「未给年份」到「已发生」,口径在走,日期在漂
Hassabis 2023–2026 「we could be just a few years, maybe within a decade away」(2023-05)→「50% chance by in the next 5 years…by 2030」(2025-07)→「five to ten years」(2026-01,二手转述) 窗口 2023–2030 持续收窄并稳定在「5–10 年」
Amodei 2024–2026 「I think it could come as early as 2026, though there are also ways it could take much longer」(2024-10)→「as little as 1–2 years away」(2026-01) 2026 窗口正在被检验
LeCun 2023–2026(反方) 「there is no such thing as AGI」;「it’s certainly not next year like our friend Elon has said」 其「不是几年内」判断截至 2026 暂时成立;「LLM 是死路」是未定路线之争

3.2 经典四条:全部未兑现,且一条被本人否认

Newell & Simon 1957/1958: 演讲于 1957 年 11 月 ORSA 年会,正式发表于 1958 年《Operations Research》6(1):1–10。[一手逐字] 四条预言逐字(含常被省略的限定语「unless the rules bar it from competition」与第四条约「most theories in psychology will take the form of computer programs」)。兑现:国际象棋冠军实际拖到 1997 年 Deep Blue,约 40 年而非 10 年;「发现并证明重要新数学定理」至今无公认案例(2024 年 IMO 金牌级解题仍是「解题」不是「发现定理」)。来源:CMU Simon 档案扫描件

Simon 1960/1965: 考古勘误四: 通行版本「1965 年《The Shape of Automation》」的出处最早是 1960 年《The New Science of Management Decision》p.38。[一手逐字]「Technologically, as I have argued earlier, machines will be capable, within twenty years, of doing any work that a man can do.」来源:Quote Investigator 扫描件核对

Minsky 1967: [一手逐字]「Within a generation, I am convinced, few compartments of intellect will remain outside the machine’s realm—the problems of creating ‘artificial intelligence’ will be substantially solved.」通行版常删去「I am convinced」前半句。来源:Quote Investigator全书扫描

Minsky 1970(最重要的一条勘误): 《Life》1970-11-20 记者 Brad Darrach 的报道确有此句:「In from three to eight years we will have a machine with the general intelligence of an average human being. I mean a machine that will be able to read Shakespeare, grease a car, play office politics, tell a joke, have a fight.」[一手逐字] 但 Minsky 本人 1996-06 邮件否认数字:[一手逐字]「Yes, this did appear in a story in Life magazine. McCarthy, Fredkin, and I considered suing the author, but decided that there’d be no point to it…I have no idea where he got the 3 to 8 years. A plausible interpretation is that we said something like ‘in 2000 or 2050’, and the writer or editor changed ‘decades’ into ‘years.’」来源:Fun_People 档案(Minsky 1996 邮件)。同访谈「keep us as pets」句系误传小说《Colossus》剧情,Minsky 亦否认。

理论整合/推断(记分卡第一课): 经典四条里,最常被引用的 Minsky 1970 恰恰是最不可靠的一条——它登在杂志上、被当事人否认、且数字可能被编辑改动。预言史的第一条纪律是:先核实预言本身,再谈兑现

3.3 Lighthill 1973 与 AI 冬天

[一手逐字]「Most workers in AI research and in related fields confess to a pronounced feeling of disappointment in what has been achieved in the past twenty-five years…In no part of the field have the discoveries made so far produced the major impact that was then promised.」来源:Lighthill 报告全文 PDF。后果:英国 SRC 削减 AI 资助,引爆第一次 AI 冬天;McCarthy 撰文反驳(存档)。

理论整合/推断(记分卡第二课): 1973 年的「失望判词」在当时是准确的评估,却押错了「通用机器人注定失败」的长期判断——2010s 深度强化学习与机器人的再繁荣翻了案。上一轮的落空,不等于下一轮也落空——这正是「AGI 预测」两端(神化与虚无)都要记住的对称教训。

3.4 Musk:唯一可记账的「顺延序列」

时间 原话(转引级别注明) 对应目标年
2023-07 「digital superintelligence in roughly the five- or six-year timeframe」(可靠转引) 2028–29
2023-12 世界「less than three years away from AGI」(可靠转引) 2026
2024-04 「If you define AGI as smarter than the smartest human, I think it’s probably next year, within two years」(可靠转引) 2025–26
2025-12-17 AGI「within the next couple of years, and maybe as soon as 2026」(内部全员会,二手转述) 2026
2026-01-06 「I think we’ll hit AGI next year in ’26」(一手音频+转录) 2026

我们的断言: 这是全库记分卡里唯一一组可以记账的「顺延序列」——2024 年 4 月的「明年/两年内」与 2025 年 12 月的「2026」在 2025 年末即告落空(媒体直接以「AGI Prediction Just Got a Two-Year Extension」为题),但每一轮落空后都会给出新的短窗口。序列的价值不是证明「Musk 必错」,而是证明:在没有完工判据的情况下,年份预言是免检产品——落空不需要付出任何记账成本。来源:Gizmodo 2025-12-17转录稿

3.5 2023–2026:历史上第一次「从业者+研究者」同押短窗口

Hassabis 收敛: 2023-05「maybe within a decade away」(二手逐字转引)→2025-07-25 Lex Fridman「My estimate is sort of 50% chance by in the next 5 years. So, you know, by 2030, let’s say.」(一手视频)→2026-01-15 CNBC「five to ten years」(二手转述)→2026-06/07 斯坦福 GSB「maybe 2030, plus or minus a year」(可靠转引)。来源:Lex Fridman #462PCMag/AxiosZME Science。另:DeepMind 首席 AGI 科学家 Shane Legg 给 2028 年 50%「minimal AGI」(同源转引,需亲核)。

Altman 的漂移: 2023-02-24 官方博客(日期勘误:非 2023-01-24)只给「soon or far in the future」[一手逐字];2025-01-05《Reflections》(主笔亲核全文):[一手逐字]「We are now confident we know how to build AGI as we have traditionally understood it. We believe that, in 2025, we may see the first AI agents ‘join the workforce’ and materially change the output of companies」;2025-08-08 CNBC(主笔经转录核对):[一手逐字]「I think it’s not a super useful term」;2025-12-18 Big Technology 播客(转录级):「we’ve got it wrong with AGI, we never defined that…my proposal is that we agree that you know AGI kinda went whooshing by. It didn’t change the world that much…we built AGIs.」来源:Planning for AGI and beyondReflectionsCNBC 2025-08-11Windows Central 转录

Amodei 的「powerful AI」: 主笔 2026-08-01 亲核《Machines of Loving Grace》全文:[一手逐字]「What powerful AI (I dislike the term AGI) will look like, and when (or if) it will arrive…I think it could come as early as 2026, though there are also ways it could take much longer」;能力定义:「it is smarter than a Nobel Prize winner across most relevant fields」;「We could summarize this as a ‘country of geniuses in a datacenter’」;2026-01《The Adolescence of Technology》:「powerful AI could be as little as 1–2 years away」(主笔亲核原文检索命中)。来源:Machines of Loving GraceThe Adolescence of Technology

理论整合/推断(记分卡第三课): 2023–2026 是预言史上前所未有的局面:三家里有两家(OpenAI/Anthropic)从「不给年份」转向「给 2–5 年窗口」,研究者调查与预测市场同步把中位数压向 2030 年代。经典预言全落空的历史并不自动适用于本轮——但正因为本轮是「从业者自证」,它的记账标准必须更严:给窗口的人同时要给判据,否则窗口就是免检产品

四、专家调查:中位数、口径与「一年前移 13 年」

4.1 Grace 系列的三轮公开数字(全部回到原文逐字)

前置勘误: ① 流传的「2022 年 1 月版 arXiv:2201.01662」不存在——该 ID 是一篇数学论文;2022 轮于 2022-06-12 至 08-03 施测,2022-08-03/04 发布(官方页与 JAIR 论文均注明)。② 系列没有《…Year Two》副标题,正式名称为 Expert Survey on Progress in AI(ESPAI)。③ 2022 与 2023 两轮合并在一个论文里发表:arXiv:2401.02843(2024-01)→ JAIR 84:9(2025)。

HLMI 问题原文(2016–2023 逐字一致): [一手逐字]「Say we have ‘high-level machine intelligence’ when unaided machines can accomplish every task better and more cheaply than human workers. Ignore aspects of tasks for which being a human is intrinsically advantageous, e.g. being accepted as a jury member. Think feasibility, not adoption.」来源:arXiv:1705.08807arXiv:2401.02843

三轮中位数(HLMI):

轮次 样本 50% 概率年份 10% 概率年份 来源
2016 N=352(HLMI 题 n=259) 2061(45 年后) 2025(9 年后) arXiv:1705.08807
2022 N=738(HLMI 题 n=461) 2059(37 年;官方页原文),重算后 2060 2029 AI Impacts 2022 官方页JAIR 2024 论文脚注
2023 N=2,778(HLMI 题 n=1,714) 2047 2027 arXiv:2401.02843 摘要,主笔亲核

[一手逐字](2023 轮摘要,主笔亲核):「the chance of unaided machines outperforming humans in every possible task was estimated at 10% by 2027, and 50% by 2047. The latter estimate is 13 years earlier than that reached in a similar survey we conducted only one year earlier.」「the chance of all human occupations becoming fully automatable was forecast to reach 10% by 2037, and 50% as late as 2116 (compared to 2164 in the 2022 survey).」

考古勘误五: 「2022 版 25% 在 2037 前、10% 在 2026 前」不见于任何官方材料——2022 官方 10% 值是 2029;「2026」疑为 2016 轮的二手误传,「2037」疑为 2016 轮 FAOL 20% 百分位的误传。引用时不得使用。

4.2 调查的自我限定与框架效应

框架效应(同卷两倍差): 2022 年官方博客:[一手逐字]「Looking at just the people we asked for years, the aggregate forecast is 29 years, whereas it is 46 years for those asked for probabilities.」2023 年论文(方向相反,更极端):[一手逐字]「the year with a 50% chance of HLMI from participants answering in the fixed-year frame (34 years) was twice as far into the future as that for participants answering in the fixed-probability frame (17 years).」来源:AI Impacts 2022 博客arXiv:2401.02843

HLMI vs FAOL 差 70 年: 2023 论文把「机器在所有任务上胜过人类」(HLMI,2047)与「所有职业完全自动化」(FAOL,2116)分开——同一批受访者,两个问法,中位数差 69 年。AI Impacts 内部方法学批评(Tom Adamczewski, 2024-12-16):[一手逐字] 现行 gamma 分布+概率均值使尾部严重失真(FAOL 90% 值高达 2843 年);HLMI/FAOL 分开报道「left room for selective interpretations」;Google Scholar 引用中 7 篇有 6 篇只引 HLMI 数字;用中位数聚合重算合并 p20/p50/p80 = 2048/2073/2103。来源:AI Impacts 方法学批评

专家自认分歧: 2023 论文:[一手逐字] 元问题(n=671):「44% said ‘Not much,’ 46% said ‘A moderate amount,’ and 10% said ‘A lot’」(自认与典型 AI 研究者存在分歧);图 3 注明「the thinner ‘confidence interval’ in 2023…is due to our increased confidence about the average respondents’ views due to a larger sample size, not respondents’ predictions converging」——置信区间收窄≠预测收敛。来源:arXiv:2401.02843

2024 轮未发布: Grace 本人 2025-10-31 FAQ:「2024 (results coming soon!)」——截至 2026-08-01,最新公开数字仍是 2023 轮(2047)。来源:AI Impacts FAQ

4.3 对照调查

  • UCL《Visions, values, voices》(2024 夏施测,2025-03 发布): N=4,260 完成;51% 同意「AGI is inevitable」;无时间线中位数题。来源:ZenodoNature 报道
  • AAAI 2025 主席小组调查(475 人): 77% 优先「可接受风险收益比」而非直接追求 AGI;76% 认为「scaling 当前方法到 AGI」不可能或很不可能;82% 主张 AGI 公共所有。来源:AAAI 2025 报告 PDF
  • 2026 存在性安全峰会会前调查(59 名 AI 安全领袖): 50% 概率 AGI 的中位数 2033(均值 2034)。来源:GreaterWrong 汇总
  • FRI LEAP 第 8 波(2026-06-02): 专家中位数——2100 年前「多数专家认定 AGI 存在」概率 80%、条件中位数 2050(超级预测者 2047);METR 8 小时任务 80% 成功率中位数 2030。来源:Forecasting Research LEAP Wave 8

理论整合/推断: 把四条线并排看——学术作者调查(HLMI 2047)、安全领袖(2033)、超级预测者(2047–2033)、高管(2026–2030)——口径从宽到窄,中位数从 2047 走到 2026。这不是「谁对谁错」的数据,而是「同一问题在不同人群、不同定义、不同赌注下的回答差出一个时代」的证据。

五、预测市场:口径、定价与游走

5.1 Metaculus 主问题的四条硬指标(口径审计核心)

Metaculus Q5121「When will the first general AI system be devised, tested, and publicly announced?」(开于 2020-08-23,解决年 2080)的解决标准不是「AGI」通义,而是单系统须同时满足四条硬指标:[一手逐字] ①「Able to reliably pass a 2-hour, adversarial Turing test during which the participants can send text, images, and audio files」;②「Has general robotic capabilities, of the type able to autonomously…satisfactorily assemble a (or the equivalent of a) circa-2021 Ferrari 312 T4 1:8 scale automobile model」;③「High competency at a diverse fields of expertise, as measured by achieving at least 90% mean accuracy across all tasks in the Q&A dataset developed by Dan Hendrycks et al.」;④「Able to get top-1 strict accuracy of at least 90.0% on interview-level problems found in the APPS benchmark」;以及「unified」定义与三种解决途径(直接演示/开发者声明/委员会多数裁决)。来源:Metaculus Q5121

我们的断言: 「Metaculus 预测 AGI 在 2033」与「OpenAI 预测 AGI 在 2026」不是同一件事——前者定义里含组装 1:8 法拉利模型的机器人、对抗性图灵测试与 APPS 90% 三道门槛,后者的「AGI」是 Altman 自己都承认「never defined that」的词。口径跳在预测市场内部同样成立。

5.2 价格史:从「五十年后」到「十年内」再到回调

时点 平台/口径 数字
2020 Metaculus(社区中位) 「median of 50 years away」
2022-07-01 Metaculus(Q4815 已解决值) 中位 2030-06-02
2024-12 Metaculus 社区预测 25% by 2027、50% by 2031
2025 初 Metaculus(strong AGI 中位) 2031-07
2026-02 Metaculus 社区预测 25% by 2029、50% by 2033
2026 初 Metaculus(strong AGI 中位) 2033-11(较 2025 初后移 2 年 4 个月)
2026-08-01(今日快照) Metaculus Q5121 中位约 2032–2033(动态定价,时点值)
2026-08-01(API 实时) Manifold「AGI before 2030?」 35.0%
2026-08-01(API 实时) Polymarket「OpenAI 宣布 AGI 在 2027 前」 11%

来源:80,000 Hours(2025-03,含 2026-02 更新)80,000 Hours 播客文(2026-03-13)Metaculus Q4815Manifold(时点 API)Polymarket(时点 API)

我们的断言: 2025 年内的游走最能说明「定价不是科学概率」:2025 初 strong AGI 中位 2031-07 → 2026 初 2033-11(一年后移两年半,80,000 Hours 称之为「大爆炸式后移」)→ 2026-02 又「180° 反弹变短」。同一批平台、同一批问题、同一定义,定价随 2025 年模型发布节奏(Gemini 3/Opus 4.5/GPT-5 的成色)来回摆动——这是市场情绪的有效读数,不是世界状态的测量

5.3 2025 年 AGI 宣称类市场的兑现记录(记分卡)

市场 定价 结果
Metaculus「A major AI lab claim in 2025 that they have developed AGI?」 社区 0% No(2026-01-06 解决)
Polymarket「OpenAI announces it has achieved AGI in 2024?」 —($907K 成交) No
Polymarket「OpenAI announces it has achieved AGI in 2025?」 2025-02 峰值约 27% No($337K 成交)
Polymarket「Anyone win the ARC Prize Grand Prize?」 No(无人达到 85%,已定于 2025-12-31 解决)
Metaculus「AI get 85% on the ARC benchmark?」(ARC-AGI-1) 达成(2025-11-18 12:00 UTC 解决,同期 Gemini 3 Deep Think 87.5% 报道)

来源:Metaculus Q30923Polymarket 2024Polymarket 2025Polymarket ARCMetaculus Q28665

理论整合/推断(记分卡第四课): 2025 年是「AGI 宣称」的零兑现年:大厂宣称 No、ARC 85% 门槛 No、唯一 Yes 的是 ARC-AGI-1 的 85%(且是带效率约束之外的商业模型 Deep Think)。「2025 年 AGI 已实现」类叙事在当年即被市场结算为 No——这是 AGI 预言史上第一次有可审计的、公开的记账。

5.4 「2025 AI Challenge」勘误

考古勘误六: 通行说法「Metaculus 与 Arc Prize 合作举办 2025 AI Challenge」不实——ARC Prize 2025 由 ARC Prize Foundation 与 Kaggle 合作(85% 门槛、$700K Grand Prize、2025-03-26 开赛、2025-11-03 结束),Metaculus 只有社区成员自发创建的相关问题。Metaculus 官方同期项目是「AI Forecasting Benchmark Series」。来源:ARC Prize 2025 条款结果公告

5.5 「预测市场的 AGI 定价被当科学概率」的批评

  • Ben Landau-Taylor《Probability Is Not A Substitute For Reasoning》(2023):[一手逐字]「Expressing yourself in terms of probabilities does not absolve you of the necessity of having reasons for things」「slapping unjustified numbers on raw ignorance does not actually make you less ignorant.」来源:原文
  • Nuño Sempere(2023):[一手逐字]「forecasting seems to work best when working with clear definitions. And the fact that this is expensive to do makes the topic of AI a bit of a bad fit for forecasting.」来源:原文
  • Evan Harper《The Least Interesting Prediction In The World》(2026-01):[一手逐字]「One should be very cautious about interpreting Metaculus forecasts, especially historical ones. The resolution criteria must be read carefully and skeptically」;并记录 Metaculus 内部观点:「once you’ve really carefully defined a question…whether that thing is 70% likely or 80% likely, nobody cares.」来源:原文
  • RAND 报告的结构性观察:[一手逐字]「Because narrower targets are easier to achieve, this structural constraint likely biases prediction market timelines shorter than those from expert surveys that use more-expansive definitions.」来源:RAND RRA4692-1

六、合同与监管:AGI 的官方生命

6.1 微软-OpenAI 协议:AGI 的合同化与死亡

谱系:

  1. 2019-07-22:首提「pre-AGI」概念:「we intend to license some of our pre-AGI technologies, with Microsoft becoming our preferred partner for commercializing them.」[一手逐字] 来源:OpenAI 公告
  2. 2023 年(保密协议):AGI 被定义为能产生约 $100B 利润的系统,触发「AGI 条款」(IP 与 Azure 独占权豁免)。该定义从未公开——由 The Information 2024-12 披露,经 TIME/TechCrunch 转述(需亲核,二手转述):[多源交叉]「AGI is reportedly defined as being achieved when an AI system is capable of generating the maximum total profits to which its earliest investors are entitled: a figure that currently sits at $100 billion」。来源:TIME 2025-01-08
  3. 2025-10-28:AGI 宣布须经独立专家小组核实:[一手逐字]「Once AGI is declared by OpenAI, that declaration will now be verified by an independent expert panel」「Microsoft’s IP rights to research…will remain until either the expert panel verifies AGI or through 2030, whichever is first」。来源:OpenAI 官方博客
  4. 2026-02-27:「AGI definition and processes are unchanged.」[一手逐字] 来源:OpenAI 官方博客
  5. 2026-04-27:AGI 条款废止。 [一手逐字] 微软声明:「Microsoft will continue to have a license to OpenAI IP for models and products through 2032. Microsoft’s license will now be non-exclusive」「Revenue share payments from OpenAI to Microsoft continue through 2030, independent of OpenAI’s technology progress」。来源:OpenAI 官方博客微软官方博客

我们的断言: AGI 作为合同触发词在 2026-04-27 死亡——被固定日期(2030/2032)取代。这是「AGI 定义」从未有过的唯一一次完全实体化(利润口径),也是它唯一一次有确切的死亡时间。合同把它从愿景翻译成会计,会计又把它翻译回日期。

6.2 上市公司年报词频(主笔 2026-08-01 自行下载核验)

  • Microsoft 10-K FY2026(主笔下载 EDGAR 主文档、词边界检索):AGI 出现 0 次;OpenAI 提及 30 次(业务描述/竞争/股权投资/关联方收入 $24.1B),风险因素无 AGI 条款内容。
  • Alphabet 10-K FY2024/FY2025AGI 出现 0 次(两版均 0,含大小写不敏感)。
  • 来源:MSFT FY2026 10-K(EDGAR)GOOG FY2025 10-K(EDGAR)

理论整合/推断: 与 L5 篇(五份年报 Level 5 0 次)同构:越是被公众讨论的术语,在正式文件里越罕见。两家市值最大的 AI 押注者,年报正文一次都不用「AGI」——他们的正式语言是「AI」「machine learning」「models」,AGI 活在采访、博客与已废止的合同里。

6.3 美国联邦文件

  • EO 14179《Removing Barriers to American Leadership in AI》(2025-01-23)(主笔下载 govinfo 全文、词边界检索):AGI 出现 0 次;「AI」定义引用成文法 15 U.S.C. 9401(3)。来源:govinfo 全文
  • NIST AI 100-3《The Language of Trustworthy AI》词汇表:「AGI」「general intelligence」「superintelligence」「human-level」全部 0 次。来源:NIST 技术出版物
  • EU-US TTC 术语表第二版(2025-01-24):「AGI」0 次。来源:NIST 托管 PDF

我们的断言: 截至 2026-08-01,美国联邦官方文件对 AGI 的定义次数为零——唯一的「官方定义」来自一家私营公司自己的章程。一个被全世界媒体每天讨论的词,在监管语言里不存在。

七、虚无侧:批评者的真论点与无证主张

7.1 LeCun:最系统的反方

代表性原话(逐字,均附来源):

  • 「(0) there is no such thing as AGI. Reaching ‘Human Level AI’ may be a useful goal, but even humans are specialized.」(2022-05 推文,二传)
  • 「an LLM is basically an off-ramp, a distraction, a dead end」(2024-04 伦敦,TNW 转引)
  • 「LLMs are great, they’re useful, we should invest in them…They are not a path to human-level intelligence. They’re just not.」(2025-11,Business Insider)
  • 「You have all those people bloviating about AGI in a year or two. Just completely delusional, just complete delusion.」(2025-12 The Information Bottleneck 播客,the-decoder 报道;同日 Hassabis 公开回击:「confusing general intelligence with universal intelligence」「Brains are the most exquisite and complex phenomena we know of」)
  • 「The claims that somehow by just scaling up LLMs, we’re going to reach super human intelligence, that is simply not going to happen.」(2026-07-02 BBC)

来源:the-decoder 2022TNW 2024-04-10Business Insider 2025-11-17the-decoder 2025-12-22BBC 2026-07-02

我们的断言: LeCun 的批评分两层:①「AGI 现在还没到」——成立(截至 2026 无公认 AGI);②「LLM 路线到不了、必须换架构(JEPA/世界模型)」——无证主张:他与 Hassabis 的公开对骂(「complete BS」vs「extremely general brains」)双方都没有判决性证据。虚无侧的论证强度不高于神化侧:双方都在押注未验证的路线

7.2 Marcus 与 Bender

  • Marcus 2022-03《Deep Learning Is Hitting a Wall》(Nautilus):[一手逐字]「there are serious holes in the scaling argument. To begin with, the measures that have scaled have not captured what we desperately need to improve: genuine comprehension.」;2025-08-07(GPT-5 发布日):[一手逐字]「GPT-5 is obviously not AGI.」「’AGI 2027′ seems more and more remote by the day.」来源:NautilusSubstack
  • Bender 等《Stochastic Parrots》(FAccT 2021,考古勘误七:非微软研究院成果):[一手逐字]「an LM is a system for haphazardly stitching together sequences of linguistic forms it has observed in its vast training data, according to probabilistic information about how they combine, but without any reference to meaning: a stochastic parrot.」来源:ACM DL。注:论文没有逐字「没有证据表明语言模型理解」句(需亲核引文要用原文)。

7.3 「AGI 是宗教」与「移动靶」的学术化

  • Ben Recht(UC Berkeley,2025-04):[一手逐字]「The religious sect in question here is AGI-theism, those who believe that they can create superintelligence」「Just because they had a big success, it doesn’t mean their religion is correct」。来源:argmin.net
  • 《On the (Im)possibility of AGI…》(arXiv:2601.17335,2026-01):[一手逐字]「The absence of a shared definition is not a mere terminological nuisance. It prevents falsifiable claims, encourages rhetorical goalpost shifts, and makes it difficult to separate engineering progress from statements that are, in effect, metaphysical.」
  • SciAm 2025-11-20《Every AI Breakthrough Shifts the Goalposts》:[一手逐字]「the target we’ve set for AI intelligence has continually moved」;引 Hofstadter:人类在机器攻克后把任务降级为「mere mechanical abilities」。来源:SciAm

理论整合/推断: 「移动靶」描述既有实证(HLE/ARC 门槛年年后移)也有理论化(arXiv:2601.17335 的「prevents falsifiable claims」)。它是本篇判据账的核心证据:AGI 不是「判据在场而口径缺席」(固态电池),也不是「没人朝终点报进度」(L5)——而是终点线本身从未被权威地画下来过。

八、能力刻度的反面证据:可测的进展与未达的门槛

防虚无锚(能力轨迹真实可测):

刻度 2023 2025–2026 门槛
ARC-AGI-1 纯 LLM 约 0–8% Gemini 3 Deep Think 87.5%(2025-11,报道级) 100%(人类)
ARC-AGI-2 发布(2025-03)纯 LLM 0% 榜首 24.03%;refinement 54%($31/task) 85%(ARC Prize 门槛,无人达到)
HLE 发布(2024)<10% 2025-08 GPT-5 64.7%(模型卡) 无官方及格线
IMO 2024 银牌 2025 金牌(35/42) 金牌线

来源:ARC Prize 结果分析HLEDeepMind IMO 2025 公告

我们的断言: 能力刻度上的进展是真实的(ARC-AGI-2 从 0%→54%、HLE 从 <10%→64.7%、IMO 金牌),而每一个「AGI 级」门槛都没有被官方口径达成(85% 未达、HLE 无门槛)。这正是「有真有空」的样板:进展真、到达未立。「AGI 永远不来」派(虚无)与「AGI 已到/必到」派(神化)各执一端,而刻度本身只说一句话:逼近在发生,到达未发生。

九、双向裁决与落点

9.1 四向裁决

  1. 「AGI 在 X 年到来」未立: 每一个带年份的预言都缺两样东西——统一的定义与可审计的判据;已落空的经典预言(Simon/Minsky/Musk 序列)证明「年份免检产品」的机制;2023–2026 高管窗口正在被检验,但同口径可核。
  2. 「AGI 已实现」未立: 2025 年所有「AGI 宣称」类市场全部解决 No;ARC 85% 门槛无人达到;OpenAI 自己把 AGI 宣布外包给独立专家小组(且该机制从未被触发)。
  3. 「AGI 是纯炒作、永远不会来」同样未立: 能力轨迹真实可测且高速推进;专家中位数(2047)与安全领袖(2033)即使打对折也不是「永远」;虚无侧自己的路线判断(LeCun 反 LLM)同样无证。
  4. 「专家调查/预测市场给出了科学概率」未立: 框架效应差一倍、口径差 69–70 年、一年前移 13 年、时点价随事件游走、造尺人(AI Impacts)自己批评自己的聚合方法——这些数字是带口径的条件语句,不是世界状态的测量。

9.2 落点

灵魂句:AGI 从来不缺日期,缺的是终点线;而每一条被画出来的终点线,都是按画它的人的尺子刻的。

与 L5 篇、固态电池篇、抗衰路线图篇同族的最后一块拼图:那三篇的病分别是「边界由被测方自己写」「判据在场而口径缺席」「完工判据不存在/写在水位上」;本篇的病更深一层——同一个词在六个口径之间自由滑动,每个口径都自带一条不同的终点线,而没有任何一份权威文件把「哪一个口径算数」定下来。1950 年图灵把问题换成游戏的那一刻起,「AGI 何时到来」这个问题就从来没有一个可回答的形式;2026 年它的回答形式仍然取决于你先回答「你说的 AGI 是哪一个」。

十、自指与诚实空位

10.1 自指三刀

  • 本库自己也报年份: 本库的索引用「机制裁决第 N 篇」编号,报进度的方式与预测市场报「50% by 2033」同类——若「报进度须带完工判据」的纪律成立,本库的「篇数」同样不是能力刻度。本篇随文记下这条欠账。
  • 「判据账比日期账更根本」这把新尺第一次使用: 与固态电池篇的「判据在场而口径缺席」同族但不同病,声明适用范围未知,欢迎反例。
  • 本库对 AI 能力篇目的态度不对称: 本库已有 LLM 推理/ICL/Agent/Scaling 四篇 AI 能力审计,它们的结论(能力真·机制分层)与「AGI 未立」并不冲突——本篇不重审它们,也不借它们的结论反向给 AGI 判死刑。

10.2 诚实空位

  1. Metaculus 时点价: Q5121 今日中位数约 2032–2033 为动态渲染快照(Metaculus API 已改认证制),非精确值;历史游走数字来自 80,000 Hours 与 RAND 的公开记录。
  2. 「$100B 利润定义」: 2023 年保密协议内容为二手转述(TIME/The Information),未取回协议原件——不以其逐字承重。
  3. Amodei 2026-02 Dwarkesh「50% 1–3 年、90% 10 年」: 第三方转录摘要,未逐字核到音频。
  4. Musk 2024-12「AGI by 2025」: 仅聚合站转引,未找到一手文字稿。
  5. Minsky 1982「hardest science」句: 仅 Ullman 2005 转引,未见 1982 年一手出处。
  6. Grace 2024 轮结果: 截至 2026-08-01 未发布,「results coming soon」。
  7. Kurzweil 2024「可能保守」自评: 二手转引(Science Friday 转录),未核音频。
  8. 最大的结构性弱点: 本篇无法对「2026–2027 高管窗口」做事前审计——它正在发生;本篇只把账本立好(每句预言带日期、原文、场合),留给将来的读者打勾。

十一、来源清单

11.1 编号来源(分四档核验:✓ 一手亲核 / ◐ 可靠转引或存档 / ○ 二手转述 / ⚠ 未核实)

# 来源 核验
1 Turing 1950, Mind LIX(236)(umbc 抄录)
2 Gubrud 1997(Wayback 快照)
3 Goertzel MIRI 访谈 2013
4 AGI-08 会议报告(AI Magazine) ✓(Wiley 反爬,开篇句经 Goertzel MIRI 访谈交叉)
5 Introducing OpenAI 2015(Wayback)
6 OpenAI Charter(主笔亲核)
7 Chollet 2019, arXiv:1911.01547
8 Levels of AGI, arXiv:2311.02462
9 Bloomberg: OpenAI Sets Levels(2024-07-11) ◐(存档)
10 HLE 公告(CAIS/Scale)
11 HLE Nature 论文 s41586-025-09962-4
12 ARC Prize 2025 条款
13 ARC-AGI-2 发布说明
14 ARC Prize 2025 结果分析(2025-12-05) ✓(主笔亲核)
15 Hendrycks et al. 2025《A Definition of AGI》, arXiv:2510.18212
16 Newell & Simon 1958(CMU 档案)
17 Quote Investigator: Simon 1960
18 Quote Investigator: Minsky 1967
19 Minsky 1970 Life 报道与 1996 否认(Fun_People)
20 Lighthill 1973 报告全文
21 Gizmodo: Musk 2025-12-17
22 Musk 转录稿 2026-01
23 OpenAI: Planning for AGI and beyond(2023-02-24)
24 Altman: Reflections(主笔亲核全文)
25 CNBC: Altman AGI 一词无用(2025-08-11)
26 Windows Central: Altman whooshing by 转录
27 Amodei: Machines of Loving Grace(主笔亲核)
28 Amodei: The Adolescence of Technology
29 Hassabis: Lex Fridman #462(2025-07-25) ✓(视频)
30 Grace et al. 2016/2018, arXiv:1705.08807
31 AI Impacts 2022 官方页
32 Grace et al. 2023, arXiv:2401.02843(主笔亲核摘要)
33 JAIR 84:9 (2025)
34 AI Impacts 方法学批评(2024-12-16)
35 AI Impacts FAQ(2024 轮未发布)
36 UCL Visions values voices(Zenodo)
37 AAAI 2025 主席小组报告
38 2026 存在性安全峰会会前调查
39 FRI LEAP Wave 8(2026-06-02)
40 Metaculus Q5121(含解决标准) ✓(标准)/◐(时点价)
41 Metaculus Q4815(已解决 2030-06)
42 80,000 Hours 2025-03(含 2026-02 更新)
43 80,000 Hours 播客文 2026-03-13
44 Manifold AGI before 2030(时点 API) ✓(时点)
45 Polymarket OpenAI AGI before 2027(时点 API) ✓(时点)
46 Metaculus Q30923(2025 宣称 No)
47 Polymarket OpenAI AGI 2025(No)
48 Polymarket ARC Grand Prize(No)
49 Metaculus Q28665(ARC 85% 达成 2025-11-18)
50 Landau-Taylor: Probability Is Not A Substitute(2023)
51 Sempere: Hurdles of forecasting AI(2023)
52 Harper: The Least Interesting Prediction(2026-01)
53 RAND RRA4692-1
54 OpenAI-Microsoft 2019-07-22
55 OpenAI-Microsoft 2025-10-28
56 OpenAI-Microsoft 2026-02-27
57 OpenAI-Microsoft 2026-04-27
58 微软博客 2026-04-27
59 MSFT FY2026 10-K(EDGAR,主笔词频核验)
60 GOOG FY2025 10-K(EDGAR)
61 EO 14179(govinfo,主笔词频核验)
62 NIST AI 100-3 词汇表
63 EU-US TTC 术语表第二版
64 LeCun: the-decoder 2022
65 LeCun: TNW 2024-04-10
66 LeCun: Business Insider 2025-11-17
67 LeCun vs Hassabis(the-decoder 2025-12-22)
68 LeCun: BBC 2026-07-02
69 Marcus: Deep Learning Is Hitting a Wall(Nautilus)
70 Marcus: GPT-5 hot take(2025-08-07)
71 Bender et al. Stochastic Parrots(ACM)
72 Recht: AGI-theism(2025-04)
73 arXiv:2601.17335(2026-01)
74 SciAm: 移动靶(2025-11-20)
75 TIME: Altman superintelligence(2025-01-08)
76 DeepMind IMO 2025 公告
77 技术奇点篇(分界参考)
78 预测市场篇(分界参考)
79 L5 篇(技术时间表家族分界)
80 固态电池篇(判据账分界)

11.2 核验统计

  • 80 条编号来源;✓ 一手亲核 62 / ◐ 可靠转引或存档 15 / ○ 二手转述 0 / ⚠ 未核实 3(诚实空位 2/5/7 对应第 26/22/29 条外的转录级材料,已标注)。
  • 主笔 2026-08-01 亲自抓取核验:OpenAI Charter 全文、arXiv:2401.02843 摘要、ARC Prize 2025 结果公告、Amodei《Machines of Loving Grace》关键段、Altman《Reflections》全文、MSFT FY2026 10-K 词边界检索(AGI=0/OpenAI=30)、EO 14179 govinfo 全文(AGI=0)。

十二、后续问题

  1. ARC-AGI-3(2026 年初已发布) 的 85% 门槛是否在 2026 年被打破——若开源系统达成,升级「能力刻度」判断。
  2. Grace 2024 轮结果(”results coming soon”)的中位数与百分位——若 2024 轮继续前移,验证「从业者自证」加速度;若后移,验证 2025 年「大爆炸式后移」。
  3. 2026–2027 高管窗口(Amodei「1–2 年」、Musk「2026」、Hassabis「2030±1」)的兑现——本篇已把账本立好。
  4. Metaculus Q5121 的解决标准(法拉利模型/对抗图灵测试/APPS 90%)哪一条最先被满足——这是四个可独立审计的子预言。
  5. 「AGI 条款」废止后的继任者:2030/2032 固定日期是否成为新一轮合同的时间表锚。
  6. 若美国或欧盟监管开始定义 AGI(NIST/EU AI Act 后续),本篇「监管语言零出现」的裁决需更新。

关联笔记

  • 技术奇点篇(概念本体分界):docs/research/2026/2026-07-03-technological-singularity-intelligence-explosion-stress-test.md
  • Scaling 撞墙篇(能力轨迹分界):docs/research/2026/2026-07-23-scaling-wall-pretraining-data-test-time-compute-stress-test.md
  • 预测市场篇(价格≠概率复用):docs/research/2026/2026-07-20-prediction-markets-wisdom-vs-liquidity-stress-test.md
  • L5 篇 / 固态电池篇 / 抗衰路线图篇(技术时间表家族):docs/research/2026/ 下 2026-07-27 的 -l5-autonomous-driving--solid-state-battery--aging-reversal-roadmap- 三篇
  • AI 就业篇(TFP 分界):docs/research/2026/2026-07-21-ai-labor-market-substitution-complement-stress-test.md