The New York Times wants OpenAI and Microsoft to pay for training data | TechCrunch
The New York Times has filed a major copyright lawsuit against OpenAI and Microsoft, alleging that their generative AI models were trained on millions of the newspaper’s articles without consent. This legal action highlights the intense conflict between content creators and AI developers, who argue that web scraping falls under fair use, while publishers insist on compensation for the unauthorized use of their proprietary journalism. The suit seeks to force the destruction of affected training data and demands significant damages, marking one of the most significant challenges to the current AI training methodologies. Beyond financial losses, The Times argues that these AI systems are creating direct competitors that erode their business model by regurgitating content and bypassing paywalls. This case underscores the broader open data dilemma regarding the ownership, licensing, and ethical sourcing of training datasets. It raises critical questions about whether the vast corpora of publicly available web data should remain free for AI consumption or if creators deserve rights to control and monetize how their work is used to build commercial AI products. This dispute is crucial for the open data community as it tests the limits of intellectual property in the age of machine learning. The outcome will likely define the future of data accessibility for AI development, potentially forcing a shift toward licensed data markets or stricter technical barriers. If successful, such lawsuits could reshape how open data is aggregated, limiting the free flow of information necessary for training robust models and impacting the broader ecosystem of open-source and commercial AI innovation.
Source: techcrunch.comPublished on 2023-12-28
Related news
- 'The New York Times' demanda a OpenAI y Microsoft por infracción de derechos de autor para entrenar la IA
- The NYT demanda a OpenAI y Microsoft por infracción a derechos de autor, al utilizar sus artículos para entrenar chatbots
- ‘Data poisoning’: How artists are fighting back against Artificial Intelligence image generators