YOU ARE VIEWING ONE ITEM FROM THE AICRIER FEED

Microsoft, OpenAI filings admit scraping theft

AICrier tracks AI developer news across Product Hunt, GitHub, Hacker News, YouTube, X, arXiv, and more. This page keeps the article you opened front and center while giving you a path into the live feed.

// WHAT AICRIER DOES

7+

TRACKED FEEDS

24/7

SCRAPED FEED

Short summaries, external links, screenshots, relevance scoring, tags, and featured picks for AI builders.

Microsoft, OpenAI filings admit scraping theft
OPEN LINK ↗
// 2h agoPOLICY REGULATION

Microsoft, OpenAI filings admit scraping theft

Newly unsealed filings in The New York Times' copyright lawsuit against OpenAI and Microsoft expose internal communications showing executives at both companies were acutely aware that their AI training practices harmed publishers and skirted legal boundaries. Internal records reveal Microsoft leadership described AI scraping as 'the largest theft of labor in human history' while telemetry showed Copilot slashed publisher referrals by up to 93%, alongside OpenAI staff celebrating paywall bypasses and acknowledging ChatGPT as substitutive of journalism.

// ANALYSIS

These unredacted admissions land like a precision strike on the AI industry's fair use defense by supplying direct, executive-level evidence of market substitution, intentional circumvention, and bad faith.

  • **Gutting the Fair Use Defense:** The central pillar of fair use under U.S. copyright law weighs market harm and substitution; internal records showing OpenAI leadership defining ChatGPT as "largely substitutive" alongside Microsoft tracking a 93% drop in referral traffic dismantle claims of purely transformative use.
  • **Evidence of Willful Infringement:** Records showing researchers celebrating paywall bypasses with OpenAI's president ("ah nice") and deliberately removing copyright notices to hide origins severely undermine good-faith protections and expose the firms to statutory damages.
  • **The Publisher Doom Loop:** Microsoft's internal acknowledgment that LLMs are destroying the economic viability of their own "content supply chain" articulates the broader structural crisis facing foundation models: consuming and starving the human knowledge sources they depend on to remain relevant.
  • **Impending Legal Reckoning:** Despite recent amicus support from the federal government favoring fair use for AI training, explicit paper trails demonstrating systemic piracy and paywall evasion make high-value settlements or restrictive judicial rulings far more inevitable.
// TAGS
microsoftopenaicopyrightintellectual-propertyartificial-intelligencethe-new-york-timesfair-useweb-scrapingcopilotchatgpt

DISCOVERED

2h ago

2026-09-18

PUBLISHED

4h ago

2026-09-18

RELEVANCE

9/ 10

AUTHOR

pluc