OpenAI and Microsoft Knew Web Scraping Was a Doom Loop
Unsealed court documents in the New York Times lawsuit call it the largest theft of labor in human history.
Unsealed court documents in the New York Times lawsuit against OpenAI and Microsoft show the companies had internal warnings before they acted. Their own documentation flagged a doom loop that would damage the open web, described their training data scraping as the largest theft of labor in human history, and said it made a complete mockery of the idea of fair use. They built anyway.
The Verge reports that the most pointed quotes come from Microsoft's Director of Applied Science, Brent Hecht. Microsoft has since tried to distance itself, with spokesperson Alex Haurek stating the comments reflect one employee's views. That distancing move is notable: internal documentation and a named director-level source are harder to dismiss than a leaked memo.
The operating question now is liability. If internal documentation shows companies knew the harm and proceeded, fair use defenses weaken. Publishers and regulators should watch how courts treat the doom loop framing, because that language came from inside the house. This is not a case of an external critic raising alarms. It is the defendants' own record.
Analysis
Internal awareness of harm versus public claims of good faith: that gap is where liability lives. Who pays here is no longer abstract.
Research this with your AI
Copy the research prompt into your AI assistant to see how this story affects you.
Show the prompt
I just read this AI news story and want to understand it in my own context. Title: OpenAI and Microsoft Knew Web Scraping Was a Doom Loop Summary: Newly unsealed court documents in the New York Times case against OpenAI and Microsoft show both companies internally warned their practices would create a web-damaging doom loop. Microsoft's Director of Applied Science, Brent Hecht, called the scraping the largest theft of labor in human history. Category: Industry Source: The Verge, https://www.theverge.com/ai-artificial-intelligence/997633/openai-microsoft-chatgpt-ai-new-york-times-doom-loop-theft-google-zero Using my own history and context, help me understand: 1. What is the core development and why does it matter? 2. Who are the major players involved and what are their motivations? 3. How does this fit into the broader AI landscape right now? 4. How does this apply to my own work, and what should I do or watch next? Be specific and plain spoken.
Newsletter
The day's AI stories, with the editor's take, in one email.
Free. Unsubscribe in one click.