Seattle Times and Newsday sue OpenAI and Microsoft for infringement
The Seattle Times and Newsday have filed lawsuits against OpenAI and Microsoft, accusing copyright infringement for using their journalism as training data for AI models without permission. The plaintiffs say the training processes relied on articles produced by their reporters, and that passages from their reporting appear in responses generated by the models when users pose questions. The actions underscore ongoing tensions between publishers seeking protection for their work and the fast-changing world of generative AI.
In the complaints, the outlets assert that their news coverage was scraped, indexed, and incorporated into training datasets used to build AI systems. They also argue that the technology sometimes reproduces fragments of their reporting in search results or answer snippets, potentially diluting the value of professional journalism and undermining licensing arrangements publishers rely on to fund coverage.
The Verge AI notes that these cases add to a growing set of lawsuits against OpenAI and other AI developers, as rights holders seek remedies for perceived misuses of content. The lawsuits against OpenAI and Microsoft come amid a broader debate about fair use, licensing, and the responsibilities of technology firms to secure rights before leveraging journalism for training data.
Legal experts say outcomes could shape how publishers license content and how AI firms source data for training in the coming years.
- The suits name OpenAI and Microsoft as defendants, according to the filings.
- The plaintiffs claim their journalism was used to train AI models without permission and that passages from reporting appear in AI responses.
- Observers say the cases are part of a wave of copyright challenges facing AI companies.
- Industry analysts caution that the litigation could influence licensing norms and data sourcing practices for training data.
- Experts emphasize that the legal landscape around AI training data remains unsettled and rapidly evolving.
For readers tracking the AI copyright debate, these lawsuits illustrate a central tension: the push to unleash powerful AI capabilities while ensuring content creators are compensated and rights are respected. As The Verge AI reports, publishers are increasingly scrutinizing how training data is collected and used—and what recourse they have when content is repurposed by AI models without consent.
The outcome of these cases could have wide-reaching consequences for publishers, AI developers, and end users who rely on AI-powered tools. If courts require more stringent licensing or stricter controls on data use, the business models for both AI startups and established tech firms could shift in meaningful ways.
