Court Documents Show Microsoft and OpenAI Anticipated Web 'Doom Loop'
Unsealed filings in the New York Times lawsuit reveal internal warnings about AI chatbots destroying content supply chains and constituting 'largest theft of labor in human history.'
Internal warnings went unheeded
Newly unsealed court documents from the New York Times' copyright lawsuit against OpenAI and Microsoft reveal that employees at both companies explicitly warned leadership about the destructive impact their AI chatbots would have on web publishers and content creators.
The filings, first reported by The Verge, include statements from Microsoft's Director of Applied Science Brent Hecht characterizing the companies' data scraping practices as the "largest theft of labor in human history" and saying their fair use defense made "a complete mockery" of the legal doctrine.
Microsoft has attempted to distance itself from Hecht's views. Company spokesperson Alex Haurek stated that the comments "reflect one employee's individual perspective" and "do not represent the company's views." Another Microsoft executive described Hecht's role as intentionally adversarial, employed to bring "asymmetrical, futuristic, and academic points of view."
The supply chain paradox
Perhaps most striking is an internal Microsoft document acknowledging that "Our AI content strategy has started a 'doom loop' that will hurt the performance of our models and the entire web at the same time." The document notes it is "highly unusual that an end-product threatens the economic foundations of its essential suppliers."
Microsoft CEO Satya Nadella acknowledged in the filings that chatbots have essentially replaced traditional search, removing the need for users to visit original sources. OpenAI's Head of ChatGPT is quoted saying that once users receive an answer from the chatbot, there is "no good reason to click" through to source material.
The companies' own analyses attributed significant drops in referral traffic for publishers directly to AI summaries, with speculation that search referrals may have declined as much as 60 percent.
Awareness of copyright violations
The documents reveal OpenAI was aware of ChatGPT's tendency to reproduce copyrighted material verbatim. Internal communications acknowledged that GPT-4 "memorized a ton of data and therefore will be insanely good at regurgitation," despite recognizing that preventing such memorization was important to "minimize copyright violations."
The filing cites multiple examples of ChatGPT outputting extended passages directly from articles in the Times, Mercury News, Denver Post, and other publications.
An OpenAI representative admitted being "unaware" of any effort to detect or remove paywalled content from training data, despite Nadella later stating that "anything that is paywalled should be licensed."
Why it matters
These revelations demonstrate that the disruption to digital publishing wasn't an unforeseen consequence but a predicted outcome that both companies chose to accept. The documents suggest leadership prioritized potential revenue—OpenAI cofounder Greg Brockman referenced "gazillions" of dollars—over concerns about damaging the content ecosystem their models depend upon. For business leaders in media and content-dependent industries, this underscores the need for proactive strategies to protect intellectual property and revenue streams as AI adoption accelerates.
Commercial motivations acknowledged
Microsoft's own internal documents admitted that "almost no one intended for they [sic] content they created to be used in this fashion, nor are they compensated for its use." OpenAI Policy Director Jack Clark recognized the companies were "creating systems that substitute for the labor of the people that define the 'culture' of society."
The 92-page filing was first reported by The Verge as part of ongoing litigation between the New York Times and the AI companies.
This is an original analysis by the Omega editorial team. Source reporting: The Verge.
Want systems like this working for your business?
Book a Call
