cd /news/ai-policy/microsoft-staff-asked-if-ai-scraping… · home topics ai-policy article
[ARTICLE · art-133673] src=decrypt.co ↗ pub= topic=ai-policy verified=true sentiment=↓ negative

Microsoft Staff Asked If AI Scraping Was 'Largest Theft of Labor in Human History'

Court documents unsealed Thursday in the New York Times' copyright suit against OpenAI and Microsoft show Microsoft employees in 2023 debated whether AI models scraping news articles amounted to "the largest theft of labor in human history," with one internal memo warning the public would see models "hoovering up" their work as "an astonishing theft of unprecedented proportions." Microsoft attributed the memos to Brent Hecht, a director of applied science, saying they do not represent company views, while CEO Satya Nadella testified that "anything that is paywalled should be licensed by anyone who wants to use it." The Times suit, filed in late 2023 and joined by eleven other publishers, has already forced OpenAI to preserve 20 million ChatGPT conversation logs, and Judge Sidney Stein of the Southern District of New York is weighing summary judgment motions.

by read3 min views2 publishedSep 18, 2026
Microsoft Staff Asked If AI Scraping Was 'Largest Theft of Labor in Human History'
Image: Decrypt (auto-discovered)

In brief

  • Documents unsealed Thursday come from the New York Times' copyright suit against OpenAI and Microsoft.
  • A 2023 Microsoft memo warned the public would see models "hoovering up" their work as "theft of unprecedented proportions."
  • Microsoft CEO Satya Nadella testified that anything paywalled should be licensed by whoever wants to use it.

Microsoft employees discussed whether OpenAI's use of news articles amounted to "the largest theft of labor in human history," and could set off a "doom loop" that degraded the models they were building, according to court documents unsealed on Thursday and reported by the New York Times.

The filings come from the suit the Times brought against both companies in late 2023, since joined by eleven other publishers. OpenAI has contested the claims throughout, and the case has already forced it to preserve 20 million ChatGPT conversation logs. Judge Sidney Stein of the Southern District of New York is weighing summary judgment motions, and documents are being unsealed as he does.

"Millions of people around the world will soon consider large models 'hoovering up' all their work to be an astonishing theft of unprecedented proportions," one internal Microsoft document from 2023 said. Large AI models "are a product that destroys its supply chain," the same author wrote.

Microsoft says those memos were written by Brent Hecht, a director of applied science who also held a Northwestern University post, and do not represent company views. He was not a decision maker and was employed to "present divergent and asymmetric perspectives," it said in a filing.

“Leaving a mess on the carpet” #

Satya Nadella, Microsoft's chief executive, testified that "anything that is paywalled should be licensed by anyone who wants to use it," and said that had he known OpenAI was training on paywalled content he would have exercised Microsoft's right to make it retrain its models. A spokesman said he "spoke to broad principles" regarding how people find and consume information.

At OpenAI, a staffer told president Greg Brockman about building "hack” to bypass the NYT paywall. Brockman replied: "ah nice."

Nick Turley, who ran the ChatGPT team, wrote in June 2023 that AI posed an "existential threat" to publishers, and in February 2024 that AI products "will get more and more substitutive as they get better." Elsewhere he wrote that AI "products are largely substitutive, period."

An OpenAI engineer wrote in February 2023 that "no matter how prominently we show the links, users won't click," a finding that cuts against the argument chatbots send traffic back to publishers.

In a 2020 memo to Brockman and Sam Altman, then-policy director Jack Clark warned the company was "creating systems that substitute for the labor of the people that define the 'culture' of society," and would "become the symbol of how Silicon Valley is thoughtlessly stepping into other parts of life and leaving a mess on the carpet." Clark left to co-found Anthropic, which referred a request for comment from the Times to OpenAI.

Both Microsoft and OpenAI argue the training was fair use, transforming articles into new work rather than substituting for the originals. "The world can see what OpenAI and Microsoft thought all along about the fairness of their own behavior," said Steven Lieberman, who represents the New York Daily News and seven other papers.

The Times, a plaintiff in the case, declined to comment to its own reporters, who said OpenAI did not respond to requests for comment.

── more in #ai-policy 4 stories · sorted by recency
── more on @microsoft 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/microsoft-staff-aske…] indexed:0 read:3min 2026-09-18 ·