News

New York Times Alleges Microsoft and OpenAI Were Aware That Using News Content Constituted Copyright Infringement

The New York Times has intensified its legal battle against major AI companies by alleging that both Microsoft and OpenAI understood that using news content to train their artificial intelligence systems would constitute copyright infringement. This development marks a significant escalation in the ongoing dispute over how AI companies source their training data and whether the use of copyrighted material without explicit licensing constitutes fair use or unlawful copying.

The lawsuit centers on claims that large language models, which power popular AI products like ChatGPT and Copilot, were trained on vast quantities of copyrighted news articles, books, and other creative works without the permission of their creators or publishers. The New York Times argues that this practice is not merely an incidental by-product of building powerful AI systems but a deliberate choice that the companies knew raised serious legal concerns.

This case is part of a broader wave of litigation brought by content creators, authors, and publishers against AI companies over training data practices. Industry observers suggest that the outcome could have far-reaching implications for how AI systems are developed and what constitutes acceptable use of copyrighted material in the development of machine learning models.

Sources