OpenAI, Microsoft warned of web’s ‘doom loop’

▼ Summary
– Unsealed court documents in the New York Times lawsuit against OpenAI and Microsoft reveal internal warnings about the negative impact of AI on the web.
– Microsoft employee Brent Hecht described data scraping for AI training as the largest theft of labor and criticized fair use defenses as a mockery.
– Microsoft attempted to distance itself from Hecht’s views, labeling them as individual academic perspectives rather than official company stances.
– Internal communications indicate that Satya Nadella acknowledged a ‘doom loop’ where AI chatbots replace search engines, threatening content suppliers.
– The filings highlight concerns from various executives about existential threats to publishers and potential massive financial gains for AI companies.
Unsealed court documents in the ongoing legal battle between The New York Times and tech giants OpenAI and Microsoft have revealed alarming internal warnings about the impact of artificial intelligence on the web. These filings suggest that both companies were aware they were initiating a “doom loop” that would severely damage the internet ecosystem, describing their data scraping practices as the “largest theft of labor in human history” and arguing that their defense strategy made a None“complete mockery of the idea of fair use.”
Many of the most critical statements originated from Brent Hecht, Microsoft’s Director of Applied Science. Although the company has attempted to distance itself from his views, the records show he was not alone in his concerns. Microsoft spokesperson Alex Haurek told The Verge that “These comments reflect one employee’s individual perspective, are not a legal analysis, and do not represent the company’s views.” In a separate filing, Jordan Usdan, GM for Data Strategy and Ops at Microsoft AI, characterized Hecht’s role as adversarial. He stated that Hecht “holds divergent, academic, and forward-looking views about how data ecosystems for AI should operate and is employed at Microsoft to bring asymmetrical, futuristic, and academic points of view … nor is he someone who speaks for Microsoft specifically as to his theoretical views on AI’s potential effect on content creators.”
Despite these attempts to isolate Hecht’s opinions, the broader narrative in the 92-page document indicates that the predicted negative outcomes have already materialized. The filings include insights from top executives such as Satya Nadella and Sam Altman, highlighting a consistent theme: the companies recognized the existential threat their models posed to publishers but proceeded regardless.
Acknowledging the Existential Threat
The introduction of the court filing features quotes from Hecht and OpenAI’s Head of ChatGPT, presumably Nick Turley, which suggest the companies understood they posed an “existential threat” to news organizations like The New York Times. Hecht explicitly labeled the harvesting of data by ChatGPT and Copilot as the “largest theft of labor in human history.” Furthermore, he argued that Microsoft’s legal stance undermined fundamental copyright principles, stating it makes a “complete mockery of the idea of ‘fair use.’”
Internal communications reveal that leadership was fully cognizant of the circular danger inherent in their business model. Satya Nadella acknowledged that chatbots had effectively replaced traditional search engines, reducing the need for users to visit original sources. More damningly, an internal Microsoft document warned that “Our AI content strategy has started a ‘doom loop’ that will hurt the performance of our models and the entire web at the same time: It is highly unusual that an end-product threatens the economic foundations of its essential suppliers, but that is the situation we have created for our LLM business with respect to its ‘content supply chain.’”
Profit Motives and Technical Realities
While public statements often emphasized altruistic goals, internal discussions pointed toward significant financial ambitions. OpenAI cofounder Greg Brockman focused on the potential earnings, referring to “gazillions” of dollars rather than ethical considerations regarding data usage. This profit-driven approach extended to technical implementation details. Despite Nadella later asserting that “anything that be licensed,” an OpenAI representative admitted during the proceedings that he was “unaware” of any systematic efforts to detect or remove paywalled content from training datasets.
The documents also expose OpenAI’s awareness of its models’ tendency to reproduce copyrighted material verbatim. Employees acknowledged that while preventing memorization was important to “minimize copyright violations,” GPT-4 “memorized a ton of data and therefore will be insanely good at regurgitation.” The filing cites numerous examples where ChatGPT outputted long strings of text directly from articles published by The New York Times, Mercury News, The Denver Post, LifeHacker, and Eurogamer.
Impact on Content Creators and Traffic
Microsoft’s internal assessments showed a clear understanding of how its scraping practices would be perceived by creators. The company admitted that “almost no one intended for they [sic] content they created to be used in this fashion, nor are they compensated for its use.” OpenAI Policy Director Jack Clark recognized the broader cultural implications, noting the creation of systems that “substitute for the labor of the people that define the ‘culture’ of society.” Internal memos described ChatGPT as “the modern newsstand,” while Nick Turley noted that once users received answers from the chatbot, there was “no good reason to click” on links to the original source.
This dynamic has led to a collapse in referral traffic for major publishers. OpenAI’s own media and economic experts attributed the decline directly to AI summaries like Google’s AI Overviews, speculating that search referrals may have dropped by as much as 60 percent. Microsoft conceded that “LLMs are a product that destroys its own supply chain” because the AI serves as a substitute for the very content it needs to train.
Although Microsoft spokesperson Haurek cautioned that “Satya’s testimony and Microsoft’s position in this case are perfectly consistent. He spoke to broad principles and changes underway in how people find and consume information. Those observations should not be confused with conclusions about copyright questions before the Court, which Microsoft addresses in its filings,” the unsealed documents paint a different picture. They suggest that both companies knew they were causing irreparable harm to the publishing industry and the millions it employs, yet continued forward in pursuit of massive profits, treating the resulting “doom loop” as an acceptable cost.
(Source: The Verge)



