{"id":82479,"date":"2026-09-20T07:09:08","date_gmt":"2026-09-20T11:09:08","guid":{"rendered":"https:\/\/overcentral.com\/en\/?p=82479"},"modified":"2026-09-20T07:09:08","modified_gmt":"2026-09-20T11:09:08","slug":"microsoft-and-openai-call-genai-largest-theft-in-history","status":"publish","type":"post","link":"https:\/\/overcentral.com\/en\/microsoft-and-openai-call-genai-largest-theft-in-history\/","title":{"rendered":"Microsoft and OpenAI Call GenAI Largest Theft in History"},"content":{"rendered":"<p>The tech industry\u2019s relationship with the creative and knowledge sectors has long been described in terms of disruption, transformation, and, more recently, extraction. But a newly unredacted legal filing has stripped away the euphemisms, revealing that leaders at Microsoft and OpenAI themselves have privately characterized their generative artificial intelligence models with a word far more damning: theft. The documents, unsealed in a New York federal court, contain internal communications where executives describe the technology as, in one stark phrasing, the \u201clargest theft of labor in human history.\u201d<\/p>\n<h2>Internal Language Exposes a Raw Assessment of GenAI<\/h2>\n<p>The filings, part of an ongoing lawsuit, offer an uncommonly candid window into how the architects of the most powerful large language models (LLMs) view their own work. OpenAI\u2019s policy director, Jack Clark, wrote that the company was \u201ccreating systems that substitute for the labor of the people that define the \u2018culture\u2019 of society.\u201d This is not a statement from a detached critic, but from an insider describing the product\u2019s core economic logic.<\/p>\n<p>Microsoft\u2019s internal policy documents were even more blunt. They asserted that generative AI could \u201csignificantly disrupt the employment of the very people who generated the data on which the foundation model was trained.\u201d The document continues with a chillingly precise description of the business model: \u201cLLMs are a product that destroys its supply chain.\u201d In other words, the raw material of these models\u2014the articles, books, code, and art created by human beings\u2014is not simply used; it is consumed in a process that, at scale, eliminates the need for the creators themselves.<\/p>\n<p>This admission from Microsoft and OpenAI confirms what many journalists, authors, and artists have long suspected: that the economic promise of generative AI is built directly on the devaluation of human-generated content. The internal language reframes the public debate. It is no longer a question of hypothetical disruption. The companies themselves have acknowledged that their products, by design, render their own source of training data economically unviable.<\/p>\n<h3>The Destroyed Supply Chain: What the Filing Reveals About Journalism<\/h3>\n<p>The legal documents provide specific evidence regarding the threat posed to news publishing. OpenAI\u2019s own internal assessment labeled the company an \u201cexistential threat\u201d to news organizations. This acknowledgment is particularly striking given the public, multi-million dollar partnerships OpenAI has pursued with major publishers. The unredacted filings suggest these deals may be less about building a sustainable ecosystem for journalism and more about securing access to high-quality training data before the supply chain collapses entirely.<\/p>\n<p>A deposition from an OpenAI software engineer further clarified the dynamics of web traffic, the lifeblood of digital publishing. The engineer testified that \u201cno matter how prominently we show the links, users won\u2019t click.\u201d This single statement dismantles the industry\u2019s prevailing narrative that AI-generated search results and summaries will drive discovery and traffic to original sources. The engineer\u2019s testimony indicates that a generation of users, when given a complete answer generated by an LLM, has no incentive to visit the original article. The link, even when prominently displayed, becomes a vestigial organ.<\/p>\n<p>This has immediate, concrete implications for anyone building a business on writing. The economic model of advertising-supported journalism relies on page views. If generative AI provides the answer without the click, the advertising revenue disappears, and the incentive to produce the original reporting vanishes.<\/p>\n<h2>How Does the \u201cTheft\u201d Actually Work at Scale?<\/h2>\n<p>To understand the scale of the problem, one must understand the mechanics of training a large language model. These models do not \u201cread\u201d articles the way a human does. They process vast quantities of text\u2014billions of words scraped from the public internet, including copyrighted news articles, books, and academic papers. During training, the model identifies statistical patterns in how words are used together. It does not store a copy of an article, but it learns the patterns that article contains.<\/p>\n<p>The critical issue is that this process occurs without the permission of the copyright holders, and until recently, without compensation. When the model is later asked a question, it uses these learned patterns to generate a statistically plausible response. For factual questions, this response often closely resembles the phrasing and information contained in the specific articles it was trained on. The filing reveals that Microsoft and OpenAI recognized this not as transformative fair use, but as the substitution of human labor\u2014a direct replacement of the journalist by the machine.<\/p>\n<p>The phrase \u201cdestroys its supply chain\u201d perfectly captures the vicious cycle. A human writer produces a detailed article. A company like OpenAI trains a model on that article and thousands like it. The model then generates answers that satisfy a user\u2019s query, eliminating the need for the user to visit the original article. Without traffic, the publication loses revenue. Without revenue, the publication cannot afford to employ the writer. The writer stops producing the article. The supply of high-quality, original data dries up. The model, unable to learn from new human work, becomes stagnant and increasingly prone to error as it is trained on its own previous output\u2014a process known as model collapse.<\/p>\n<h3>A Growing Ecosystem of Stochastic Gruel<\/h3>\n<p>The problems identified in the Microsoft and OpenAI filings are not limited to the actions of these two companies. A parallel ecosystem has emerged in which the technology is used to create vast quantities of low-quality content specifically designed to capture advertising revenue. A recent investigation detailed the operations of a media firm that buys established journalism outlets and systematically converts them into what one critic described as \u201clakes of stochastic gruel.\u201d<\/p>\n<p>These operations follow a playbook that exploits the very structure of the internet. They publish enormous volumes of AI-generated articles that are grammatically sound but factually empty, designed to rank in search engines. The investigation noted that this strategy enjoys \u201cenormous success at capturing reader eyes and attention through platforms that actual journalism relies on\u2014while publishing oceans of trash that would get a real journalist fired immediately.\u201d The firm\u2019s responses to questions about its practices were described as \u201cempty and contradictory platitudes about transparency, accountability, and responsible AI use.\u201d<\/p>\n<p>This is the environment that the unredacted filings illuminate. It is not a distant hypothetical. It is the current reality of the information economy. A major technology company has admitted internally that its flagship product is designed to destroy the industry that provides its foundational material, while a network of smaller firms uses the same technology to flood the digital commons with noise.<\/p>\n<h2>The Broader Cultural and Economic Implications<\/h2>\n<p>The language used by Jack Clark at OpenAI is particularly revealing. He described the systems as substituting for the labor of the people who \u201cdefine the culture of society.\u201d This moves the discussion beyond journalism and into the realm of cultural production as a whole. If the labor of writers, artists, musicians, and filmmakers is the raw material for these systems, and if those systems subsequently replace the creators, the long-term cultural consequence is a flattening of creative output.<\/p>\n<p>Humans write, paint, and compose from a combination of lived experience, emotion, and a unique historical perspective. A language model has no experience, no emotion, and no perspective. It has only a statistical map of what has already been said. If the economic incentive to produce new, original cultural work is removed because that work is immediately absorbed and devalued by an AI, the cultural conversation becomes a closed loop of recombined past data. The output becomes increasingly derivative, a mirror of a mirror with no new light being shined on it.<\/p>\n<p>The internal documents suggest that the executives building these systems are acutely aware of this dynamic. They have not stumbled upon it. They have identified it as the core feature of their product.<\/p>\n<h3>What This Means for the Individual Creator<\/h3>\n<p>For a freelance writer, a game journalist, or a novelist, the implications of these filings are deeply personal. The promises of the technology\u2019s advocates\u2014increased productivity, democratized access to information, a new era of creativity\u2014ring hollow when set against the cold calculus of the destroyed supply chain. The individual creator is no longer competing against another creator with a better idea. They are competing against a system that has ingested their entire corpus of work and can now reproduce its patterns for free.<\/p>\n<p>The path forward is uncertain. Legal battles over copyright and fair use will continue for years. New compensation models, such as collective licensing, are being proposed but face significant logistical hurdles. The most immediate and reliable strategy for an individual creator remains the cultivation of a direct relationship with an audience. Substacks, podcasts, Patreon-supported communities, and live events represent a return to a patronage model where the audience pays directly for the work, bypassing the advertising-driven platforms that are being hollowed out by AI-generated content.<\/p>\n<p>The badger may have taken the cat\u2019s dinner, but the cat\u2014plump, glossy, disdainful of the front door\u2014still has its own resources. So too must the writer. The data supply chain is under attack, but the human need for a narrative, for a perspective, for a story told from a specific point of view, has not changed.<\/p>\n<h2>A New Lens for Understanding the Technology<\/h2>\n<p>The unredacted court filings do not simply provide evidence for a legal argument. They provide a lexicon. The phrase \u201cdestroy its supply chain\u201d should replace the vague, optimistic language of \u201cdisruption\u201d in every business meeting, every boardroom, and every public relations statement. It is a more accurate descriptor of the mechanism at work.<\/p>\n<p>When next a tech executive speaks of \u201cdemocratizing creativity,\u201d one can now look at the internal documents of the largest companies in the sector. What they found was not a beautiful vision of a post-scarcity creative utopia. What they found was a cold, clear-eyed admission that the business plan is to vacuum up the work of the present to build a product that eliminates the need for the future. It is the largest theft of labor in human history, and the perpetrators have, in a moment of legal necessity, signed their own confession.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>The tech industry\u2019s relationship with the creative and knowledge sectors has long been described in terms of disruption, transformation, and, more recently, extraction. But a newly unredacted legal filing has stripped away the euphemisms, revealing that leaders at Microsoft and OpenAI themselves have privately characterized their generative artificial intelligence models with a word far more [&hellip;]<\/p>\n","protected":false},"author":7,"featured_media":0,"comment_status":"closed","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"fifu_image_url":"","fifu_image_alt":"","footnotes":""},"categories":[2],"tags":[],"class_list":["post-82479","post","type-post","status-publish","format-standard","category-videogames"],"_links":{"self":[{"href":"https:\/\/overcentral.com\/en\/wp-json\/wp\/v2\/posts\/82479","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/overcentral.com\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/overcentral.com\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/overcentral.com\/en\/wp-json\/wp\/v2\/users\/7"}],"replies":[{"embeddable":true,"href":"https:\/\/overcentral.com\/en\/wp-json\/wp\/v2\/comments?post=82479"}],"version-history":[{"count":1,"href":"https:\/\/overcentral.com\/en\/wp-json\/wp\/v2\/posts\/82479\/revisions"}],"predecessor-version":[{"id":82482,"href":"https:\/\/overcentral.com\/en\/wp-json\/wp\/v2\/posts\/82479\/revisions\/82482"}],"wp:attachment":[{"href":"https:\/\/overcentral.com\/en\/wp-json\/wp\/v2\/media?parent=82479"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/overcentral.com\/en\/wp-json\/wp\/v2\/categories?post=82479"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/overcentral.com\/en\/wp-json\/wp\/v2\/tags?post=82479"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}