Perplexity AI under scrutiny over illegal web scraping

Perplexity AI faces serious allegations for bypassing copyright protections and robots.txt directives to scrape content from news outlets, including paywalled material. This aggressive data acquisition strategy raises significant ethical and legal questions regarding intellectual property rights in the digital age. The incident highlights the tension between open information access and creator protections, a core concern for the open_data community. It underscores the urgent need for transparent data sourcing practices and strict adherence to web scraping conventions to ensure equitable treatment of content creators. As Perplexity’s valuation soars despite AWS investigations and founder controversies, this case serves as a critical cautionary tale. It demonstrates how rapid AI scaling without ethical data governance can undermine trust and sustainability, urging the open data ecosystem to prioritize responsible collection methods over unrestricted access.

Source: thehindu.com
Published on 2024-06-29