Skip to main content

Local News Outlets Sue Microsoft and OpenAI Over Unpaid AI Training

Local News Outlets Sue Microsoft and OpenAI Over Unpaid AI Training

A coalition of U.S. local news organizations, led by Mississippi-based Emmerich Media Group, has filed a lawsuit against Microsoft and OpenAI, accusing the tech giants of systematically using thousands of copyrighted news articles to train their AI models without permission or compensation. The complaint, filed in federal court, alleges that the companies generated billions in revenue from AI products while content creators "received nothing."

A Pattern of Unauthorized Use

This isn't the first time Microsoft and OpenAI have faced such accusations. The New York Times and other major outlets have already taken legal action on similar grounds. But the involvement of local media marks a significant expansion of the conflict—from big-city newsrooms to community newspapers that often serve as the only source of local information.

The plaintiffs say their articles were scraped and fed into AI training pipelines, sometimes even after being placed behind paywalls. They argue that this practice not only violates copyright law but also threatens the very existence of local journalism, which has already been battered by declining print sales and ad revenue.

At the heart of the lawsuit is the charge that Microsoft and OpenAI removed copyright management information—such as author names, publication titles, and usage terms—before using the content for training. The plaintiffs liken this to tearing the name tags off family photos. They also claim that AI systems have regurgitated their news stories almost verbatim, proving the content was deeply embedded in the models.

A Fight for Survival

For these local publishers, the stakes go beyond copyright. In many communities, they are the last remaining source of local news. If they are forced to shrink or close, the information vacuum could leave residents without critical coverage of local government, schools, and events. The lawsuit calls the tech companies' actions "the death knell of local journalism."

The plaintiffs accuse Microsoft and OpenAI of violating the U.S. Copyright Act and the Digital Millennium Copyright Act (DMCA). They also point to a double standard: while the companies allegedly ignore publishers' rights, they aggressively protect their own code and models through licensing and legal means. The complaint even notes that OpenAI has previously complained about competitors using its generated content to train models—a stance that contrasts sharply with its own alleged behavior.

The local media outlets are seeking damages and an injunction that would force Microsoft and OpenAI to remove all infringing content from their training data. As more news organizations and creators join the legal fray, the question of whether AI training data is legal—and how content should be licensed and compensated—is becoming one of the most pressing issues in the global AI industry.

Key Points

  • Lawsuit filed: Emmerich Media Group and other local publishers sue Microsoft and OpenAI for unauthorized use of thousands of articles in AI training.
  • Copyright infringement alleged: Plaintiffs claim removal of copyright notices and violation of the Copyright Act and DMCA.
  • Survival at stake: Local news outlets warn of a "death knell" for community journalism if AI continues to siphon traffic and revenue.
  • Demands: Damages and an injunction to remove all infringing content from AI training data.
  • Growing legal trend: The case adds to a wave of lawsuits over AI training data and content licensing.