Internal Admissions Surface in NYT Lawsuit
The legal battle between The New York Times and AI giants Microsoft and OpenAI has taken a dramatic turn with the unsealing of court documents containing candid, and potentially damaging, admissions from within the AI industry. The New York Times, which sued OpenAI and Microsoft in December 2023 for alleged mass copyright infringement, has submitted a new legal brief that leverages these internal statements to bolster its case. The publication is now pushing for a summary judgment, arguing that the evidence of unauthorized use of its copyrighted material is overwhelming. At the heart of the revelation is a statement attributed to a Microsoft director, who, in an internal discussion, reportedly described the practice of large-scale AI model training on web-scraped data as "the largest theft of labor in human history." This candid assessment, revealed in the legal filings, directly contradicts the public-facing narrative often presented by AI companies, which typically frame their data acquisition as a necessary and transformative process for advancing AI capabilities. The sheer scale of the data involved in training models like OpenAI's ChatGPT and Microsoft's Copilot, often scraped indiscriminately from the internet, is now being directly challenged not just by content creators but by figures within the very companies conducting the scraping. Adding to the pressure, the same legal brief highlights remarks from OpenAI's CEO, Sam Altman, who is quoted as calling ChatGPT an "existential threat" to publishers. This admission underscores a growing awareness within the AI industry of the disruptive, and potentially destructive, impact its technology has on traditional media business models. The NYT's legal team is using these internal acknowledgments to demonstrate that the companies were aware of the implications of their data practices and the potential harm to content creators, even as they continued to pursue them.The Scale of Data and Copyright Concerns
The core of The New York Times' lawsuit revolves around the allegation that OpenAI and Microsoft trained their AI models using millions of copyrighted articles published by the Times without permission or compensation. The NYT contends that this massive ingestion of its content allowed the AI models to generate outputs that are often derivative of the original works, directly competing with the Times and undermining its ability to monetize its journalism. The legal brief submitted by the Times aims to prove that the AI companies knew they were infringing copyright and proceeded regardless.
