Friday, September 18, 2026
Google search engine
HomeGadgetsOpenAI and Microsoft knew they were starting a ‘doom loop’ for the...

OpenAI and Microsoft knew they were starting a ‘doom loop’ for the web


Recently unsealed court documents in the New York Times’ case against OpenAI and Microsoft are pretty damning. The companies’ own documentation warned that it was starting a “doom loop” that would damage the web, characterized its scraping of data to train its models as the “largest theft of labor in human history,” and that it made a “complete mockery of the idea of fair use.”

Many of the most eye-catching quotes from the document come from Microsoft’s Director of Applied Science, Brent Hecht. Though, the company has tried to distance itself from Hecht’s assertions. Microsoft spokesperson Alex Haurek told The Verge that “These comments reflect one employee’s individual perspective, are not a legal analysis, and do not represent the company’s views.”

In a separate court filing, Jordan Usdan, GM for Data Strategy and Ops at Microsoft AI, characterized Hecht’s role as adversarial. He said that Hecht “holds divergent, academic, and forward-looking views about how data ecosystems for AI should operate and is employed at Microsoft to bring asymmetrical, futuristic, and academic points of view … nor is he someone who speaks for Microsoft specifically as to his theoretical views on AI’s potential effect on content creators.”

But whether or not Microsoft wants to own these comments, it’s clear that this came true. Google Zero is real! AI is eating the web!

There are plenty more wild statements in NYT’s filing from a variety of figures, including Satya Nadella, Sam Altman, and other OpenAI employees. Here are some highlights from the 92 page document.

“An astonishing theft”

This case is about, as Microsoft’s Director of Applied Science [Brent Hecht] put it, “an astonishing theft of unprecedented proportions”; SF1437, perhaps the “largest theft of labor in human history.”SF1652. Defendants repeatedly copied millions of Plaintiffs’ copyrighted articles in their entiretywithout permission to produce substitutive commercial AI products. OpenAI’s Head of ChatGPTwrote that “[p]ublishers” face an “existential threat” from those products, SF1466, which, he said,“are largely substitutive, period” and “will get more and more substitutive as they get better.”SF1473-74. Such admissions eviscerate Defendants’ “fair use” defense because substitution is“copyright’s bête noire.” Andy Warhol Foundation for the Visual Arts, Inc. v. Goldsmith, 598 U.S.508, 528 (2023). For Defendants to prevail on this defense “would,” the same Microsoft executiverecognized, arguably “make a complete mockery of the idea of ‘fair use.’” SF1450.

The introduction quotes Hecht and OpenAI’s Head of ChatGPT (presumably Nick Turley) in a way that seems to show the companies knew they posed an “existential threat” to publishers like the New York Times. Hecht calls ChatGPT and Copilot’s harvesting of data the “largest theft of labor in human history” and says that Microsoft’s defense makes a “complete mockery of the idea of ‘fair use.’”

Microsoft’s CEO Satya Nadella agreed under oath that conversing with chatbots “hassubstituted … giving you the information right there on the website on the AI platform versusneeding to go to the underlying source.” SF1432. A Microsoft document recognizes that nobodywins that contest: “Our AI content strategy has started a ‘doom loop’ that will hurt the performanceof our models and the entire web at the same time: It is highly unusual that an end-product threatensthe economic foundations of its essential suppliers, but that is the situation we have created for ourLLM business with respect to its ‘content supply chain.’”

Satya Nadella admits that chatbots have basically replaced search and removed the need to go straight to the source for info. But perhaps more damning is an internal Microsoft document that says, “Our AI content strategy has started a ‘doom loop’ that will hurt the performance of our models and the entire web at the same time: It is highly unusual that an end-product threatens the economic foundations of its essential suppliers, but that is the situation we have created for our LLM business with respect to its ‘content supply chain.’”

That’s not even a real number

Around the same time, OpenAI co-founderGreg Brockman wrote he was “deeply motivated by the gazillions” he hoped to gain bycommercializing OpenAI’s technology. SF630. Lately, it has been reported that OpenAI isplanning an IPO based on a $1 trillion valuation.

Don’t be fooled by OpenAI or Microsoft’s claims of altruistic intent. OpenAI cofounder Greg Brockman is more interested in the “gazillions” of dollars it he could potentially make through commercial AI.

Individuals within OpenAI and Microsoft ignored such issues as circumventing paywallsand violating terms of use when acquiring data. SF521-47, 790-92. For example, OpenAI’scorporate representative testified that he was unaware of “any effort to detect paywall content inits training datasets” or “to remove paywall content from its training datasets.”

Despite Nadella later being quoted as saying, “anything that is paywalled should be licensed,” An OpenAI representative admitted that he was “unaware” of any effort to detect or remove paywalled content from training data.

“Insanely good at regurgitation”

That same year, OpenAI recognized that its API “might outputexisting content verbatim.” SF945. By 2021, OpenAI considered the prevention of memorizationimportant “for fair use [compliance] and minimizing copyright violations in model output.” SF946.In June 2022, OpenAI employees acknowledged that GPT-4 would have “memorized a ton of dataand therefore will be insanely good at regurgitation.”

Internally, it seems that OpenAI was well aware of ChatGPT’s tendency to simply reproduce copyrighted material “verbatim.” Even though it acknowledged that the “prevention of memorization” was important to “minimize copyright violations,” employees admitted that GPT-4 “memorized a ton of data and therefore will be insanely good at regurgitation.”

The filing then goes on to cite several examples of ChatGPT outputting long strings of copy straight from articles in the Times, Mercury News, The Denver Post, LifeHacker, and Eurogamer in response to queries.

“‘Hoovering up’ all their work”

 As Microsoft recognized: “millions of people around the world will soon considerlarge models ‘hoovering up’ all their work to be an astonishing theft of unprecedented proportions”and admitted that “almost no one intended for content they created to be used in this fashion, norare they compensated for its use.”

Microsoft knew how its wholesale scraping of the internet would be perceived and admitted that “almost no one intended for they [sic] content they created to be used in this fashion, nor are they compensated for its use.”

A “substitute for the labor of people”

OpenAI Policy Director Jack Clark similarly wrote: “[O]ur work on AI and Creativity isgoing to increasingly lead to us creating systems that substitute for the labor of the people thatdefine the ‘culture’ of society[.]” SF1677. OpenAI internal documents characterize ChatGPT as“[t]he modern newsstand,” SF1500, and brag that ChatGPT provides “fast, timely answers…whichyou would have previously needed to go to a search engine for” including “up-to-date sportsscores, news, stock quotes, and more.”

OpenAI Policy Director Jack Clark saw the writing on the wall, saying that it was “creating systems that substitute for the labor of the people that define the ‘culture’ of society.” Internal documents described ChatGPT as “the modern newsstand.” OpenAI’s Nick Turley is later quoted as saying that once you get an answer from its chatbot, there is “no good reason to click” on a link to the source.

Destroying their own supply chain

Defendants acknowledge the predictable consequences of this design. Per Microsoft, the“[p]romise of LLMs is largely in the same information work domains from which they get theircontent... They naturally compete with their content supply chain.” SF1798. They substitute forthe “labor of the people” who produced the original content on which they were trained, including,among other things, newspapers and books. SF1452, 1677. There is a “real risk” that GenAI could“significantly disrupt[] the employment of the very people who generated the data on which thefoundation model was trained.” SF1467. “LLMs are a product that destroys its supply chain.”

Microsoft is quoted as admitting that “LLMs are a product that destroys its own supply chain” because it’s a substitute for its own training data in many cases.

OpenAI knows its killing referral traffic

Another OpenAI economic expert, Dr. Goldfarb, opined: “I find that the decline in referral trafficto The Times’s properties has been driven by a combination” of factors including “e.g., Google AIOverviews.” SF1784, 1786. Dr. Goldfarb also opined that “Google’s introduction of AI overviewsmay have depressed search referrals by 20 to 60 percent” for DNP. SF1785. Dr. Sinnreich,OpenAI’s media expert, opined based on a 2026 Reuters Institute analysis that “referral traffic topublishers from both Google Search and Google Discover has dropped considerably (from over 5billion monthly referrals via Discover to fewer than 4 billion, and from well over 3 billion viaSearch to slightly more than 2 billion) since Google introduced AI overviews.” SF1787. Dr.Sinnreich also admitted that declines in Google referrals are related to AI-generated summaries,among other things.

OpenAI’s own media and economic experts attributed the drop in referral traffic for sites like the Times directly to AI summaries like Google’s AI Overviews. They’ve speculated that search referrals may be down as much as 60 percent.

Microsoft spokesperson Haurek cautioned that “Satya’s testimony and Microsoft’s position in this case are perfectly consistent. He spoke to broad principles and changes underway in how people find and consume information. Those observations should not be confused with conclusions about copyright questions before the Court, which Microsoft addresses in its filings.”

But it seems pretty clear based on this newly unsealed document that both Microsoft and OpenAI knew they were going to irreparably harm the publishing industry, the “millions of people” it employs, and, by extension, damage their own product, but carried forward anyway in pursuit of “gazillions” of dollars — doom loop be damned.

Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.




Source link

RELATED ARTICLES

LEAVE A REPLY

Please enter your comment!
Please enter your name here

- Advertisment -
Google search engine

Most Popular

Recent Comments