Just in:
Minimal Phone 2 shifts focus with AMOLED upgrade // ExfilSquad data leaks substantiate broad breach claims // Biotechnology and AI Reshape Medicine Development Across the UAE and Beyond // NINGJI Takes Centre Stage at KLCC, Strengthening Its Position as a Benchmark for Southeast Asian Expansion Through Five Key Localization Strategies // Bora Group Posts Record 2Q26 Revenue and Strong Profits as Margins expand and Operations Resume Demand-Driven Growth // DP World sets $3 billion global expansion drive // Dubai shares retreat as earnings deepen market caution // Messi returns as Inter Miami bow out // Google Cloud sets 2027 quantum-security milestone // Oil steadies as blockade threat revives supply risk // NINGJI Takes Centre Stage at KLCC, Strengthening Its Position as a Benchmark for Southeast Asian Expansion Through Five Key Localization Strategies // “YOU BRING CHARM TO THE WORLD — The 18th Global Chinese Awards” Concludes in Beijing // Forest City Highlights Nearly 40 International Awards and Certifications // Leaderless Student Protests Pose New Challenge To National Politics // Etiqa Insurance Singapore Appoints Claudia Soh as Chief Executive Officer to Lead Next Chapter of Growth // Oman races to contain spreading tanker oil spill // XTransfer Serves Over 1 Million Enterprise Clients // Anthropic weighs $6 billion Decart deal // PayerMax Enables Last War to Integrate Rakuten Pay, Expanding Market Access to Japan // Firebird launches Armenia AI hub with vast expansion //

Perplexity’s Persistent Web Crawling Raises Ethical Concerns

Perplexity, a growing artificial intelligence company, has been repeatedly crawling websites against the wishes of content owners, prompting a wave of concern over digital ethics and user privacy. Despite multiple requests for the company to halt its web scraping activities, Perplexity continues to disregard these refusals, highlighting tensions between AI innovation and data ownership.

AI-driven companies have long been at the forefront of technological development, but their web scraping practices often remain shrouded in controversy. Web scraping, or the automated extraction of data from websites, is integral to training AI models. However, the practice has sparked debates regarding the ethicality of harvesting content from sites without permission. Perplexity, which leverages web data to fuel its language models, is now in the spotlight after it ignored clear signals from website owners asking it to stop.

Several notable websites have reported receiving persistent crawls from Perplexity’s bots, even after issuing direct requests to cease their activities. The situation escalated after Perplexity failed to respect robots. txt files, which are used to instruct web crawlers on which pages to avoid. Such disregard for the standard protocol has led to frustration within the web community, with some arguing that companies like Perplexity are exploiting the openness of the internet without considering the broader implications of their actions.

Tech industry leaders have weighed in, calling for better oversight of AI data collection practices. They argue that while AI can enhance innovation, there should be clear boundaries about the data it can use. In this case, Perplexity’s actions seem to suggest a disregard for the principles of consent and fairness in the digital landscape. The company’s approach to scraping raises important questions about the potential risks of AI’s reliance on web data and the legal grey areas it creates.

Legal experts have also warned that continuing to ignore these restrictions could lead to significant legal consequences for AI companies. The General Data Protection Regulation in the European Union, for example, offers users and companies a degree of control over how their data is collected and used. Infringing on these regulations could result in hefty fines for companies that fail to comply.

Despite these concerns, some within the AI community have defended Perplexity’s actions. Advocates argue that web scraping is an essential tool for advancing AI technology, claiming it enables the development of smarter and more efficient models. They contend that without access to vast amounts of data from the internet, AI systems would struggle to achieve the level of sophistication required to tackle complex tasks such as natural language processing and decision-making.



Notice an issue?

Arabian Post strives to deliver the most accurate and reliable information to its readers. If you believe you have identified an error or inconsistency in this article, please don't hesitate to contact our editorial team at editor[at]thearabianpost[dot]com. We are committed to promptly addressing any concerns and ensuring the highest level of journalistic integrity.


Loading next story…