Microsoft Exec Called AI Data Scraping 'Theft of Labor,' Filings Show
Newly unsealed court filings reveal Microsoft privately described AI data practices as 'theft' while the company and OpenAI scraped paywalled content.
Unsealed court documents indicate that Microsoft internally referred to the data practices used in AI development as 'theft.' This occurred while Microsoft and OpenAI were scraping content from paywalled publications, building datasets from this material, and internally warning about the potential impact on publishers.
The filings shed light on the ethical and legal considerations surrounding AI development. Internal communications from Microsoft used strong language to describe the potential consequences for publishers whose content was utilized. This situation has spurred broader discussions about copyright and the responsible acquisition and use of data by AI companies.
The close relationship between tech giant Microsoft and AI firm OpenAI is underscored by these revelations, which also highlight potential internal conflicts and concerns regarding data sourcing methods. Despite their collaboration on AI solutions, the documents suggest significant internal reservations were present.
These disclosures are expected to influence ongoing legal proceedings and discussions regarding AI regulation. Publishers and content creators are increasingly concerned about the unauthorized use of their material for training AI models, seeking clarity and compensation.