Unsealed court documents from a lawsuit led by The New York Times reveal that key executives at Microsoft and OpenAI internally voiced severe concerns regarding the legal and economic viability of scraping news content to train artificial intelligence models. Microsoft Director of Applied Science Brent Hecht called web scraping “an astonishing theft of unprecedented proportions” and suggested that widespread scraping made “a complete mockery of the idea of ‘fair use.’”
Internal communications show that OpenAI leadership recognized commercial AI products posed an “existential threat” to news publishers by replacing traditional media providers. Furthermore, Microsoft documented steep declines in click-through rates for publisher partners, ranging from 51% to 94%, creating what internal reports labeled a “doom loop” that threatens the underlying content supply chain required to maintain reliable AI outputs.
News organizations are utilizing these disclosures to dismantle fair use defenses by proving that AI search tools directly substitute publisher products. The documents also reveal internal awareness of methods used to bypass publisher paywalls, strengthening publisher claims as the parties prepare to go to trial.
Why it matters
Unsealed internal records undermine fair use defenses, increasing copyright legal exposure for LLM developers.
Traffic drops up to 94% signal escalating friction between AI search platforms and digital content creators.
Systemic paywall circumvention practices may face heightened judicial scrutiny and severe financial liabilities.
Source: arstechnica.com



