express gazette logo
The Express Gazette
Thursday, September 17, 2026

Tech Giants Knew AI Scraping Threatened Publishers, Court Filings Reveal

Internal communications from Microsoft and OpenAI staff expressed concerns about the impact of AI training on news organizations.

US Politics 2 hours ago
Tech Giants Knew AI Scraping Threatened Publishers, Court Filings Reveal

High-level staffers at Microsoft and OpenAI were aware that their use of online news content to train artificial intelligence models could pose a significant threat to the publishing industry, according to newly unsealed court documents. The revelations emerged as part of a copyright infringement lawsuit filed by several news publishers against the tech companies.

One OpenAI executive described the practice of scraping news content, which often circumvents publisher paywalls, as an "existential threat" to news organizations. A director at Microsoft similarly raised alarms, noting that the company's AI content strategy had initiated a "doom loop" that could negatively impact both its models and the broader web. He further commented on the unusual situation where an end product could jeopardize the financial stability of its essential suppliers.

The court filings, made public on Thursday, are part of a copyright infringement case initiated in December 2023 by The New York Times. The lawsuit, later consolidated with similar actions from publishers owned by Alden Global Capital, alleges that OpenAI and Microsoft used millions of news articles to train AI models like ChatGPT and Microsoft's Copilot.

Internal documents cited in the filing suggest that OpenAI viewed ChatGPT as a "modern newsstand." Discussions also revealed instances of staff finding ways to bypass publisher paywalls, with OpenAI President Greg Brockman reportedly responding positively to a researcher who discovered a method to circumvent The New York Times' paywall.

Microsoft CEO Satya Nadella testified that content behind paywalls should be licensed for AI training, and he indicated that Microsoft would require OpenAI to retrain its models if they were found to have used such content. Nadella also acknowledged that AI chatbots could become a substitute for visiting news websites directly.

The publishers are seeking a favorable ruling from a federal judge ahead of a potential trial, referencing internal documents and testimony from key figures at both companies. The tech companies, however, have argued that scraping content from news sites constitutes "fair use" under U.S. copyright law.

A Microsoft representative stated that the comments from its director reflected an individual perspective and not the company's official stance. The spokesperson reiterated Microsoft's position that its transformative uses of content are compliant with copyright law and that Copilot does not substitute for journalistic content. The Department of Justice has also filed a statement of interest in the case, supporting the argument that the companies' use of news content falls under fair use protections.


Sources