Friday, September 18, 2026
NewsWhite
Microsoft exec called AI scraping the “largest theft of labor in human history”
TECHNOLOGY

Microsoft exec called AI scraping the “largest theft of labor in human history”

September 17, 2026·Source: Ars Technica·1 views

A Microsoft executive has reportedly described the mass scraping of creative and professional work to train artificial intelligence models as the "largest theft of labor in human history," according to Ars Technica. The remark is striking in its bluntness, and more so given the source: a senior figure inside one of the companies most deeply invested in the AI systems that depend on exactly that practice.

To understand why this matters, it helps to step back and consider what AI training actually requires. Large language models and image-generation systems are built on enormous datasets scraped from the open web — text, code, images, journalism, fiction, academic writing, forum posts, and virtually every other form of human expression that has been committed to digital form over the past few decades. The companies building these systems have generally argued that such scraping constitutes fair use, that the models learn patterns rather than reproduce content directly, and that the practice is no different in principle from the way a human being absorbs and learns from the world around them. Courts in multiple jurisdictions are currently stress-testing those arguments, and the outcomes remain genuinely uncertain.

The creative and publishing industries have pushed back hard. Writers, visual artists, musicians, and news organizations have filed lawsuits, signed open letters, and negotiated — or attempted to negotiate — licensing arrangements with AI developers. The core grievance is consistent: their work was taken without permission, without credit, and without compensation, and the resulting systems now compete directly with the people whose labor made them possible. A staff photographer whose decades of images trained a generative model may find that model undercutting their livelihood. A journalist whose prose shaped an AI's writing style receives nothing while the company deploying that AI charges subscription fees. The asymmetry is stark.

What makes the Microsoft executive's reported remark so significant is not just its content but its provenance. Microsoft has poured billions of dollars into OpenAI and has integrated that company's technology throughout its product line, from the Copilot assistant embedded in Windows and Office to developer tools that now generate code automatically. Microsoft is not a bystander in the AI race; it is one of its most aggressive participants. For an executive inside that organization to characterize what the industry has done as the largest labor theft in human history is either a sign of genuine internal tension, a calculated positioning move, or some complicated mixture of both.

The likely reading is that the statement reflects real fractures forming within the technology industry itself as the legal and regulatory environment tightens. Companies that initially benefited from the permissive, ask-forgiveness-later culture of the early web are beginning to calculate their exposure more carefully. The creative-rights lawsuits now working through courts in the United States and Europe represent serious financial and reputational risk. Meanwhile, a growing number of publishers and platforms have begun demanding licensing deals before permitting their content to be used for AI training, which suggests the era of consequence-free scraping may be drawing to a close regardless of how the legal cases resolve.

For creators, the consequences of this moment cut in two directions. On one hand, the acknowledgment from inside the industry that something morally questionable happened is a form of validation that the creative community has been seeking. On the other hand, validation from an executive does not translate into compensation, and the practical machinery for retroactively paying the people whose work trained existing models does not currently exist and may never be built. The models are already trained. The datasets have already been consumed. Any remedies that emerge are more likely to govern future behavior than to address past harm.

For the broader technology industry, this suggests a period of accelerating fragmentation. Some companies may move toward licensing-first approaches to training data, either because they calculate that the legal risk of not doing so is too high or because they see a competitive advantage in being able to claim their systems were built on fairly sourced material. Others will continue to argue that scraping is legally and ethically defensible. The existence of that split, openly acknowledged even within individual companies, will make it harder for the industry to present a unified front as regulators in Brussels, Washington, and elsewhere consider how to draw new boundaries around AI development.

What to watch for next is whether the executive's reported remarks remain an isolated moment of candor or the beginning of a more public reckoning inside major AI developers. Equally important is how Microsoft responds if pressed to clarify the comment officially — whether it stands behind the characterization, distances itself, or tries to reframe it as a narrower point. The lawsuits moving through the courts will also continue to force these questions into the open on timelines that the industry cannot fully control. The remark, as reported by Ars Technica, has named something the sector has largely preferred to leave unnamed.

Originally reported by Ars Technica. Read the original article

Related Articles