Newly unsealed court documents from late 2023 reveal that a senior Microsoft executive privately described the mass scraping of copyrighted journalism to train artificial intelligence models as “the greatest labor theft in history.” According to records made public through a major copyright lawsuit, tech leadership across both Microsoft and OpenAI openly acknowledged the severe legal and economic risks of harvesting protected news content without authorization or financial compensation.
Internal Microsoft Documents Reveal Warnings Over AI Training
The controversy centers on internal discussions from 2023, when Brent Hecht, director of applied science at Microsoft, warned colleagues about the long-term fallout of using scraped media assets. According to documents published by The New York Times, Hecht wrote that millions of people would soon view the large-scale absorption of human work by generative models as an unprecedented intellectual property violation. Hecht further cautioned that these AI products threatened to dismantle the foundational supply chain of the independent press by replacing the need for readers to visit original publishing sites.
https://x.com/DiarioBitcoin/status/2101756816480702591
The New York Times Lawsuit and Publisher Coalitions
The disclosures emerged directly from an ongoing federal lawsuit initiated in late 2023, when The New York Times sued OpenAI and Microsoft for utilizing millions of copyrighted articles without acquiring licenses or paying fees. According to court filings covered by Xataka, eleven additional media organizations subsequently joined the legal action. Legal teams representing the publishers argue that major tech firms intentionally deployed aggressive web-scraping techniques, with internal messages showing executives celebrating methods designed to bypass paywalls to harvest restricted content.
OpenAI Executives Anticipated Workplace Displacement
Internal correspondence from OpenAI demonstrates that company leaders anticipated severe labor market disruptions years before releasing commercial chatbots. According to documents reported by El Periódico Digital, Jack Clark, then OpenAI’s head of policy, sent a 2020 report to Sam Altman and Greg Brockman warning that their models would directly substitute for the professionals who define societal culture. By June 2023, Nick Turley, head of ChatGPT development, characterized AI technology as an existential threat to journalism, noting plainly that generative tools operate primarily as substitutes rather than traditional discovery engines.

Microsoft CEO Addresses Paywall Breaches and Licensing
The unsealed records also highlight divergent reactions among executive leadership regarding data acquisition methods. According to reporting compiled by Xataka, Microsoft CEO Satya Nadella asserted that any content locked behind a paywall must remain strictly subject to licensing requirements. Nadella maintained that had he known OpenAI extracted and trained systems on paywalled information, he would have exercised Microsoft’s contractual rights to demand a complete retraining of the affected models.
- Mobile Data Subscriptions Surge in Kenya: 64.3 Million Users reached
- Fair City Tuesday Spoilers: Latest Plot Twists and Drama
- Missouri Governor Mike Kehoe Signs Executive Order on Flock Cameras and ALPR Data Privacy (news-usa.today)
- Ernie Accorsi, longtime NFL executive and Giants GM, dies at 84 (newsylist.com)