EGEGazette
Live
Houthis Accuse Saudi Arabia of 28 Airstrikes in 24 Hours; U.S. Intelligence Chief Visits CairoWashington Warns Citizens of Unexpected Escalation in Middle EastIran Says Strait of Hormuz Will Stay Closed Until US Meets Its ConditionsTrump faces dual setback as U.S. courts block voting and immigration restrictionsLawsuit Filed Against Trump and His Company Over Paid Early Access Service to His PostsTrump Renews Bid to Restrict Birthright Citizenship Through Curbing "Birth Tourism"Trump Cuts Camp David Vacation Short Amid Middle East Escalation WarningsTrump Cuts Short Vacation, Returns to White House as US Issues "Possible Escalation" Warning in Middle EastReporters From CNN, MS NOW and Politico Denied White House Access After Trump BanFederal Immigration Agent Shoots and Injures Man in Austin, TexasPowerful Explosions Reported in Syria's Aleppo CountrysideSix Pakistani soldiers, including two officers, killed in clash near Afghan border
technologyReported

Unsealed Filings Show Microsoft Executive Called AI Data Scraping 'Largest Theft of Labor in Human History'

Newly unredacted court documents reveal Microsoft privately described OpenAI's data practices as theft, even as both companies scraped paywalled New York Times content to build AI training datasets, according to TechCrunch.

· 2 min read · language: en
Article image
TechCrunch

Newly unsealed court filings show that a Microsoft executive privately characterized artificial intelligence data scraping as "the largest theft of labor in human history," according to a report by TechCrunch published Wednesday.

The unredacted documents, part of ongoing litigation, indicate that Microsoft internally referred to OpenAI's data collection practices as "theft" even as both companies were scraping paywalled content from The New York Times, TechCrunch reported.

According to the filings cited by TechCrunch, Microsoft and OpenAI used the scraped Times material to build datasets for training artificial intelligence systems. Internal communications reviewed as part of the unsealed records reportedly show employees at Microsoft warning that such practices could severely harm news publishers.

TechCrunch's report does not specify the identity of the Microsoft executive quoted, nor does it provide additional detail on the timing or origin of the internal remarks beyond their disclosure in the newly unredacted court filings.

The disclosures add to a broader legal and public debate over how AI companies obtain and use copyrighted material, particularly journalism, to train large language models. Publishers including The New York Times have pursued legal action against AI developers, arguing that unauthorized use of their content for AI training constitutes copyright infringement.

Microsoft and OpenAI have not been quoted directly in the TechCrunch report responding to the newly unsealed material. TechCrunch's account is based on the unredacted filings themselves, which the outlet reviewed as part of its coverage of the litigation.

Sources

EGazette summarizes reporting from multiple sources; follow the links for the originals.

Related articles

Comments

Sign in to join the conversation.

Forgot password?

No account?

Loading comments…