Google: New policy update allows scraping user info online to train AI
Google’s updated privacy policy explicitly permits the collection of publicly available online information to train AI models, marking a significant shift in how tech giants handle user data. This approach raises critical concerns about transparency and consent, especially amidst global legal challenges regarding unauthorized data usage. It highlights the growing tension between corporate AI expansion and individual data rights, suggesting that user-generated content may be exploited without clear acknowledgment. The implications for open data are profound, as this move blurs the lines between public information and proprietary training datasets. By scraping widely accessible sources, companies like Google may inadvertently discourage the free flow of data, which is essential for unbiased research and innovation. This lack of clarity undermines the principles of open data, potentially creating a fragmented landscape where access depends on corporate discretion rather than open standards. Furthermore, regulatory hurdles, such as those in the EU, underscore the difficulty of balancing AI advancement with privacy protections. As Google pushes deeper into generative AI without adequately addressing copyright or data protection concerns, the industry faces scrutiny over who truly controls digital information. This dynamic challenges the ethos of open data, urging a reevaluation of how public information is utilized in the age of artificial intelligence.
Source: medianama.comPublished on 2023-07-07