Microsoft exec called AI scraping 'the largest theft of labor in human history'
Unsealed court documents from a copyright lawsuit led by The New York Times reveal that Microsoft and OpenAI executives internally acknowledged serious concerns about their AI training practices before releasing ChatGPT and Copilot. Microsoft Director of Applied Science Brent Hecht called the scraping of news content for AI training "an astonishing theft of unprecedented proportions" and "perhaps the largest theft of labor in human history," while OpenAI's ChatGPT head Nick Turley warned that publishers would face an "existential threat" from AI products trained on news content. The documents show that both companies predicted and observed significant traffic declines for news organizations—with Microsoft recording 83-93 percent drops in click-through rates for some publishers—yet neither company chose to license content from news organizations before deploying their products. News organizations argue the internal evidence demonstrates that Microsoft and OpenAI knowingly violated copyright law and that AI-generated outputs frequently reproduce news articles verbatim, making fair use claims indefensible.
Why it matters A rare on-record admission from inside Microsoft about AI training data theft reframes the copyright fight at the center of the entire industry.
Why it made the cut: 940 points on Hacker News · Covered by 3 outlets