OpenAI and Microsoft accused of stealing data to train ChatGPT in new class-action suit
This article highlights ongoing legal battles concerning the ethics of data collection in artificial intelligence, specifically targeting OpenAI and Microsoft. The central issue is whether automated web scraping of private and copyrighted information without consent constitutes theft, fundamentally challenging the legitimacy of current AI training methods. This case underscores the tension between technological innovation and individual privacy rights. The lawsuit argues that these AI products rely on stolen personal data, including that of minors, to achieve their commercial success. Plaintiffs seek damages and the disgorgement of profits, asserting that the models would not exist without this alleged illegal harvesting. This implicates the broader open_data debate, as it questions the openness and legality of using publicly available web data when it involves personally identifiable information harvested without explicit permission. Relevant to open_data, this litigation threatens to restrict the free flow of information that fuels many open-source AI initiatives. If courts rule against scraping, it could create significant barriers to accessing public data for model training, potentially limiting the development of transparent and accessible AI technologies. It serves as a critical reminder that "open" access does not automatically equate to "open use," especially when privacy and consent are compromised.
Source: cointelegraph.comPublished on 2023-09-07
Related news
- The Guardian, New York Times, CNN figure in growing list of sites blocking OpenAI crawler
- Technology Innovation Institute Introduces World’s Most Powerful Open LLM: Falcon 180B
- Maya PH’s open-source LLM, Godzilla 2, surpasses ChatGPT in truthfulness
- El Inai puede reanudar sesiones
- ¿Cómo saber si estafadores lo tienen en la mira?