New York Times 诉讼解封文件显示 OpenAI 和 Microsoft 早已自知掀起危害整个网络的 doom loop

The Verge:AI(RSS)·2026-09-19 05:07·1小时前·Terrence O’Brien
AI 导读

New York Times 起诉 OpenAI 和 Microsoft 案中新解封的 92 页文件显示,两家公司内部早有记录承认其 AI 内容策略会形成伤害自身模型和整个网络的 doom loop。

The Verge:AI(RSS)
同事件
86AI 编辑部评分,满分 100

New York Times 诉讼解封文件显示 OpenAI 和 Microsoft 早已自知掀起危害整个网络的 doom loop

2026-09-19 05:07· 1小时前· Terrence O’Brien
AI 导读

New York Times 起诉 OpenAI 和 Microsoft 案中新解封的 92 页文件显示,两家公司内部早有记录承认其 AI 内容策略会形成伤害自身模型和整个网络的 doom loop。

正文 · AI 翻译

这些公司知道他们正在把我们推向“谷歌零”(Google Zero),却依然这么做了。

这些公司知道他们正在把我们推向“谷歌零”(Google Zero),却依然这么做了。

STKS534_AI_DOOMSDAY2_B STKS534_AI_DOOMSDAY2_B

最近解封的法庭文件,来自《纽约时报》对 OpenAI 和微软提起的诉讼案相当具有杀伤力。这些公司自己的文件警告称,它正在开启一个会损害网络的“末日循环”,将其抓取数据用于训练模型的行为描述为“人类历史上最大的劳动窃取”,并称这“彻底嘲弄了合理使用的理念”。

文件中许多最引人注目的引述来自微软应用科学总监 Brent Hecht。不过,该公司试图与 Hecht 的论断保持距离。微软发言人 Alex Haurek 对 The Verge 表示:“这些言论反映的是一名员工的个人观点,并非法律分析,也不代表公司的立场。”

在另一份法庭文件中,微软 AI 数据战略与运营总经理 Jordan Usdan 将 Hecht 的角色描述为对立性的。他表示,Hecht“对 AI 数据生态应如何运作持有分歧性的、学术性的和前瞻性的观点,受雇于微软正是为了带来非对称的、未来主义的和学术性的视角……他也不是以自己在 AI 对内容创作者潜在影响方面的理论观点代表微软发言的人。”

但无论微软是否愿意承认这些评论,很明显这一切都应验了Google Zero 是真的!AI 正在吞噬整个网络!

《纽约时报》的这份法庭文件中,还有大量来自各方人物的惊人言论,包括 Satya Nadella、Sam Altman 以及其他 OpenAI 员工。以下是这份 92 页文件中的一些亮点。

“一场令人震惊的盗窃”

This case is about, as Microsoft’s Director of Applied Science [Brent Hecht] put it, “an astonishing theft of unprecedented proportions”; SF1437, perhaps the “largest theft of labor in human history.”SF1652. Defendants repeatedly copied millions of Plaintiffs’ copyrighted articles in their entiretywithout permission to produce substitutive commercial AI products. OpenAI’s Head of ChatGPTwrote that “[p]ublishers” face an “existential threat” from those products, SF1466, which, he said,“are largely substitutive, period” and “will get more and more substitutive as they get better.”SF1473-74. Such admissions eviscerate Defendants’ “fair use” defense because substitution is“copyright’s bête noire.” Andy Warhol Foundation for the Visual Arts, Inc. v. Goldsmith, 598 U.S.508, 528 (2023). For Defendants to prevail on this defense “would,” the same Microsoft executiverecognized, arguably “make a complete mockery of the idea of ‘fair use.’” SF1450.

引言部分引用了 Hecht 和 OpenAI 的 ChatGPT 负责人(推测是 Nick Turley)的话,似乎表明这些公司知道自身对《纽约时报》这类出版商构成了“生存威胁”。Hecht 将 ChatGPT 和 Copilot 抓取数据的行为称为“人类历史上最大规模的劳动成果盗窃”,并表示微软的辩护“完全是对‘合理使用’这一理念的嘲弄”。

这是一个“末日循环”

Microsoft’s CEO Satya Nadella agreed under oath that conversing with chatbots “hassubstituted … giving you the information right there on the website on the AI platform versusneeding to go to the underlying source.” SF1432. A Microsoft document recognizes that nobodywins that contest: “Our AI content strategy has started a ‘doom loop’ that will hurt the performanceof our models and the entire web at the same time: It is highly unusual that an end-product threatensthe economic foundations of its essential suppliers, but that is the situation we have created for ourLLM business with respect to its ‘content supply chain.’”

Satya Nadella 承认,聊天机器人基本上已经取代了搜索,让人们不再需要直接访问信息源头。但或许更具杀伤力的是一份微软内部文件,其中写道:“我们的 AI 内容战略已经启动了一个‘末日循环’,它将同时损害我们模型的性能以及整个网络的生态:一个终端产品威胁到其关键供应商的经济根基,这是极不寻常的,但这正是我们为 LLM 业务在其‘内容供应链’方面所创造的处境。”

这甚至都不是一个真实的数字

Around the same time, OpenAI co-founderGreg Brockman wrote he was “deeply motivated by the gazillions” he hoped to gain bycommercializing OpenAI’s technology. SF630. Lately, it has been reported that OpenAI isplanning an IPO based on a $1 trillion valuation.

别被 OpenAI 或 Microsoft 所宣称的无私意图所蒙蔽。OpenAI 联合创始人 Greg Brockman 更感兴趣的是,他有可能通过商业化 AI 赚到的“天文数字”般的美元。

付费墙算什么

Individuals within OpenAI and Microsoft ignored such issues as circumventing paywallsand violating terms of use when acquiring data. SF521-47, 790-92. For example, OpenAI’scorporate representative testified that he was unaware of “any effort to detect paywall content inits training datasets” or “to remove paywall content from its training datasets.”

尽管后来有引述称 Nadella 说过,“任何设有付费墙的内容都应该获得授权”,但一位 OpenAI 代表承认,他“不知道”有任何旨在从训练数据中检测或移除付费墙内容的努力。

“在复述方面强得离谱”

That same year, OpenAI recognized that its API “might outputexisting content verbatim.” SF945. By 2021, OpenAI considered the prevention of memorizationimportant “for fair use [compliance] and minimizing copyright violations in model output.” SF946.In June 2022, OpenAI employees acknowledged that GPT-4 would have “memorized a ton of dataand therefore will be insanely good at regurgitation.”

在内部,OpenAI 似乎非常清楚 ChatGPT 倾向于直接“逐字”复制受版权保护的材料。尽管它承认“防止记忆化”对于“尽量减少版权侵权”很重要,但员工们承认,GPT-4“记住了大量数据,因此在复述方面会强得离谱”。

随后,该文件继续列举了若干例子,说明 ChatGPT 在回应查询时,直接输出了来自 TimesMercury NewsThe Denver PostLifeHackerEurogamer 文章中大段大段的原文内容。

“‘把他们所有的作品都吸走’”

As Microsoft recognized: “millions of people around the world will soon considerlarge models ‘hoovering up’ all their work to be an astonishing theft of unprecedented proportions”and admitted that “almost no one intended for content they created to be used in this fashion, norare they compensated for its use.”

微软知道其对互联网的大规模抓取会被如何看待,并承认“几乎没有人希望自己创作的内容被以这种方式使用,他们也没有因此获得补偿。”

“人类劳动的替代品”

OpenAI Policy Director Jack Clark similarly wrote: “[O]ur work on AI and Creativity isgoing to increasingly lead to us creating systems that substitute for the labor of the people thatdefine the ‘culture’ of society[.]” SF1677. OpenAI internal documents characterize ChatGPT as“[t]he modern newsstand,” SF1500, and brag that ChatGPT provides “fast, timely answers…whichyou would have previously needed to go to a search engine for” including “up-to-date sportsscores, news, stock quotes, and more.”

OpenAI 政策总监 Jack Clark 看出了不祥之兆,他表示这是在“创建替代人类劳动的系統,而这些人定义了社会的‘文化’。”内部文件将 ChatGPT 描述为“现代报刊亭”。OpenAI 的 Nick Turley 后来被引述说,一旦你从它的聊天机器人那里得到答案,就“没有好的理由去点击”指向来源的链接。

摧毁自己的供应链

Defendants acknowledge the predictable consequences of this design. Per Microsoft, the“[p]romise of LLMs is largely in the same information work domains from which they get theircontent... They naturally compete with their content supply chain.” SF1798. They substitute forthe “labor of the people” who produced the original content on which they were trained, including,among other things, newspapers and books. SF1452, 1677. There is a “real risk” that GenAI could“significantly disrupt[] the employment of the very people who generated the data on which thefoundation model was trained.” SF1467. “LLMs are a product that destroys its supply chain.”

微软被引述承认“LLM 是一种摧毁自身供应链的产品”,因为在许多情况下,它是自身训练数据的替代品。

OpenAI 知道自己在扼杀推荐流量

Another OpenAI economic expert, Dr. Goldfarb, opined: “I find that the decline in referral trafficto The Times’s properties has been driven by a combination” of factors including “e.g., Google AIOverviews.” SF1784, 1786. Dr. Goldfarb also opined that “Google’s introduction of AI overviewsmay have depressed search referrals by 20 to 60 percent” for DNP. SF1785. Dr. Sinnreich,OpenAI’s media expert, opined based on a 2026 Reuters Institute analysis that “referral traffic topublishers from both Google Search and Google Discover has dropped considerably (from over 5billion monthly referrals via Discover to fewer than 4 billion, and from well over 3 billion viaSearch to slightly more than 2 billion) since Google introduced AI overviews.” SF1787. Dr.Sinnreich also admitted that declines in Google referrals are related to AI-generated summaries,among other things.

OpenAI 自己的媒体和经济专家将《纽约时报》等网站推荐流量的下降直接归因于 Google 的 AI Overviews 等 AI 摘要。他们推测搜索推荐流量可能下降了多达 60%。

微软发言人 Haurek 提醒说:“Satya 的证词与微软在本案中的立场完全一致。他谈论的是广泛的原则,以及人们查找和消费信息的方式正在发生的变化。这些观察不应与法院面前关于版权问题的结论混为一谈,微软在其提交的文件中已就后者作出回应。”

但根据这份最新解封的文件,情况似乎相当清楚:微软和 OpenAI 都知道他们将不可挽回地损害出版业、出版业所雇用的“数百万人”,并进而损害他们自己的产品,却仍然为了追逐“天文数字”般的金钱而一意孤行——哪怕末日循环在所不惜。