Microsoft Exec Called AI Scraping 'Largest Theft of Labor'
Internal documents unsealed in a lawsuit reveal a Microsoft executive described AI data scraping as 'an astonishing theft of unprecedented proportions,' potentially the 'largest theft of labor in human history.'

Internal documents unsealed in a lawsuit accuse Microsoft and OpenAI of copyright infringement, alleging the companies used vast amounts of news content to train artificial intelligence models without proper authorization. The documents, revealed in a motion for summary judgment filed by news plaintiffs including The New York Times, shed light on the companies' internal views regarding the data scraping practices that fueled AI development.
A Microsoft Director of Applied Science, Brent Hecht, is quoted in the unsealed documents expressing significant concerns about the practice. He reportedly described the scraping of news content for AI training as "an astonishing theft of unprecedented proportions" and potentially "the largest theft of labor in human history." These statements appear to contradict the companies' public stance and legal arguments.
The news organizations contend that Hecht's internal warnings undermine Microsoft and OpenAI's defense that using news content for AI training constitutes fair use. In one document, Hecht allegedly suggested that the extensive scraping of news made "a complete mockery of the idea of ‘fair use.’" The lawsuit centers on whether the AI firms violated copyright laws by allegedly using copyrighted material from news outlets to develop products like ChatGPT and Copilot.
This unsealing of internal communications comes as part of an ongoing legal battle between major news publishers and AI companies over the use of copyrighted material. The outcome of this case could set significant precedents for the future of AI development and intellectual property rights in the digital age.