OpenAI ya no podrá utilizar el contenido del New York Times para entrenar a ChatGPT

The New York Times has fundamentally altered its terms of service to explicitly prohibit OpenAI from using its publications to train ChatGPT. This decisive move addresses growing ethical concerns regarding the unauthorized use of copyrighted material and privacy issues associated with automated content scraping tools like GPTBot. By restricting access to such high-quality journalistic data, the newspaper aims to protect the intellectual property and rights of its creators against unchecked artificial intelligence development. This restriction highlights the critical tension between technological innovation and legal ownership in the digital age. While the ban limits OpenAI’s direct ingestion of the newspaper’s content, it does not prevent individual users from manually inputting articles during interactions. Consequently, the challenge of data provenance and copyright compliance shifts from corporate scraping practices to user-generated inputs, complicating the enforcement of ethical boundaries in AI training datasets. This scenario is highly relevant to open data because it underscores the fragility of data accessibility when proprietary claims override public interest. As media outlets increasingly restrict access to prevent AI exploitation, the open data ecosystem risks becoming fragmented and exclusionary. This conflict forces a reevaluation of how data is shared and utilized, emphasizing the need for clear ethical frameworks that balance the open availability of information with the legitimate rights of content creators.

Source: laopinion.com
Published on 2023-08-16