Interestana
Home/News/News Outlets Sue OpenAI and Microsoft for Copyright Infringement
The Verge3 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

News Outlets Sue OpenAI and Microsoft for Copyright Infringement

The Seattle Times and Newsday have initiated legal action against OpenAI and its primary investor, Microsoft, asserting claims of copyright infringement. These two prominent news organizations allege that their journalistic content was utilized as training data for OpenAI's artificial intelligence models without proper authorization. The lawsuits, filed in the U.S. District Court for the Southern District of New York, contend that OpenAI's AI models have been trained on vast amounts of copyrighted material scraped from the internet, including articles published by The Seattle Times and Newsday. A key accusation is that OpenAI's models not only learned from this content but also frequently reproduce verbatim or near-verbatim passages from the plaintiffs' reporting in response to user prompts. This practice, the news outlets argue, constitutes a direct violation of their copyrights and deprives them of potential revenue streams derived from their journalistic work.

These legal challenges follow a pattern of similar lawsuits filed by other media organizations against AI companies. The Associated Press, for instance, previously sued OpenAI and Microsoft, alleging that the AI models were trained on millions of its copyrighted articles. The New York Times also filed a significant lawsuit against OpenAI and Microsoft in December 2023, accusing them of using millions of its articles to train AI chatbots, leading to the AI models producing content that directly competes with the newspaper. The core of these disputes revolves around the fair use doctrine and whether the large-scale ingestion and processing of copyrighted text for AI training purposes qualify as transformative use or constitute infringement. The plaintiffs are seeking damages and injunctive relief to prevent further alleged misuse of their intellectual property.

OpenAI, known for developing advanced AI models like GPT-3.5 and GPT-4, has publicly stated its commitment to respecting copyright and intellectual property rights. However, the company has also defended its data collection practices, arguing that the training of its models on publicly available web data is a necessary component of AI development. Microsoft, as a major investor and technology partner of OpenAI, is also implicated in these lawsuits due to its role in supporting and integrating OpenAI's technology into its products and services. The outcomes of these legal battles could have significant implications for the future of AI development, data licensing, and the media industry's ability to protect its content in the age of generative AI. The plaintiffs are seeking to establish legal precedents that would require AI developers to obtain explicit licenses or permissions before using copyrighted journalistic content for training their models, thereby ensuring fair compensation and control over their intellectual assets.

Original source — read the full reporting at the publisher:

Read on The Verge

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next