The New York Times has filed a lawsuit against Microsoft and OpenAI, alleging that these tech giants infringed on copyright by using millions of its articles without permission to train generative AI models. This legal action marks a critical turning point, as The Times becomes the first major U.S. news outlet to directly challenge these companies after failed commercial negotiations, seeking the destruction of unlicensed training data and substantial damages. This case is highly relevant to open data, as it highlights the urgent need for transparency and ethical standards in how public and proprietary information is harvested for machine learning. It underscores the tension between the open development of AI technologies and the protection of intellectual property, forcing a re-evaluation of data sourcing practices. Ultimately, the lawsuit establishes a precedent for holding AI developers accountable for unauthorized content usage. It signals that while some media outlets have opted for licensing deals, others are resisting exploitation, thereby pushing the industry toward more structured and legally compliant frameworks for data utilization in artificial intelligence development.

Source:
Published on 2023-12-28