Microsoft exec labels AI scraping as biggest labor theft in history
microsoft training
| Source: HN | Original article
A Microsoft executive described AI training as the largest theft of labor in human history.
Microsoft executives have been quoted in a newly unsealed legal brief as describing the mass‑scraping of online content for artificial‑intelligence training as “the largest theft of labor in human history.” The comment appears in the New York Times’ lawsuit against Microsoft and OpenAI, filed earlier this month and now made public through court documents.
The brief alleges that Microsoft’s Copilot reduced click‑through traffic to nytimes.com by roughly 93 percent and that OpenAI’s training sets contain more than 91,000 copies of publisher‑owned works. The executive’s private remark, captured in internal emails, underscores that senior staff were aware of the scale of data extraction and its impact on content creators.
The admission matters because it provides concrete evidence that the companies’ own leadership recognized the ethical and commercial risks of using unlicensed web data, yet continued to rely on it. It bolsters the Times’ claim that AI products are built on “astonishing theft” of copyrighted material and strengthens the broader argument that large‑language‑model training threatens the economics of news organisations. The revelation also dovetails with earlier reporting on September 18, when we highlighted the “doom loop” narrative linking AI development to the erosion of the web’s content ecosystem.
Going forward, the court’s next steps will be closely watched. The lawsuit could force Microsoft and OpenAI to renegotiate data‑licensing arrangements, adopt stricter provenance checks, or even alter the architecture of future models. Regulators in the EU and the United States have signalled interest in curbing unlicensed data use, and the case may become a catalyst for new legislation. Observers will also be looking for an official response from Microsoft, as well as any settlement talks that could reshape the relationship between AI developers and the publishing industry.
Sources
Back to AIPULSEN