Google updates its privacy policy to allow data scraping for AI training

Google’s recent privacy policy update explicitly permits the use of publicly available online data for training its artificial intelligence models. This clarification establishes a clear precedent that information exposed to the general public can be harvested and utilized by tech giants to develop and refine their advanced AI systems, shifting the boundary between public sharing and corporate data exploitation. This move occurs amidst growing legal and ethical controversies surrounding AI training practices. Recent lawsuits against competitors highlight concerns regarding unauthorized data scraping and violations of user privacy, while social media platforms are implementing stricter access controls to prevent service degradation caused by excessive automated data extraction. These industry-wide tensions underscore the urgent need for clearer regulations on how public information is processed. This development is critical to open data discussions as it blurs the line between open access and proprietary training resources. It raises significant questions about whether data shared openly for public benefit can be legally repurposed for commercial AI without consent. Consequently, it challenges the fundamental assumptions of the open data movement, highlighting the potential for public information to be absorbed into private, opaque machine learning models without transparency or user control.

Source: cointelegraph.com
Published on 2023-07-05