TTB White LOGO TB
  • News
  • PC & Hardware
  • Mobiles
  • Gaming
  • Electronics
  • Gadget
  • Reviews
  • How To
  • Login
  • Sign Up
Trending
Microsoft Launches Copilot Tasks Research Preview
Lenovo Teases Foldable Legion Go With 11.6 Inch Screen
How iOS 26 Prioritizes Alerts With Apple Intelligence
Mobile Game Installs Drop but Spending Climbs in 2025
Instagram Will Tell Parents When a Teen Searches for Suicide or Self Harm
Saturday, Sep 19, 2026
The Tech BasicThe Tech Basic
Font ResizerAa
Search
Follow US
Perplexity
The Tech Basic > News > Cloudflare Research Reveals Perplexity Crawlers Are Masking as Chrome
News

Cloudflare Research Reveals Perplexity Crawlers Are Masking as Chrome

Salman Akhtar
Last updated: 5 August 2025 12:38
Salman Akhtar
Share
Image Source: Yahoo! Tech
SHARE

Cloudflare’s security team announced it has found evidence that Perplexity’s web crawler is evading site owners’ requests to block AI bots. The infrastructure provider began investigating after customers reported unexpected scraping of their pages. In its August 4 technical report, Cloudflare said its researchers observed Perplexity changing its user-agent string and even its network identifiers to mimic legitimate browsers. These tactics allowed the crawler to slip past robots.txt rules and custom blocks meant to stop data harvesting.

Contents
  • How Perplexity Crawls Blocked Sites
  • Perplexity’s Response and Dispute
  • The Broader Battle Over AI Scraping
  • What Website Owners Can Do
Perplexity
Image Source: Storyboard18

How Perplexity Crawls Blocked Sites

Cloudflare explains that Perplexity’s crawler first identifies itself with a declared user-agent. When blocked, it switches to a generic “Mozilla Chrome on macOS” signature. It also moves its traffic across multiple autonomous system numbers to avoid IP-based defenses. Cloudflare used machine learning and network analysis to fingerprint the crawler’s behavior across millions of daily requests to tens of thousands of domains. In response, Cloudflare has removed Perplexity from its verified crawler list and updated its own tools to block these stealth requests.

Perplexity’s Response and Dispute

Perplexity spokesperson Jesse Dwyer dismissed the report as a “sales pitch” and insisted the screenshots showed no actual content was retrieved. In a follow-up, Dwyer claimed the bot Cloudflare identified “is not even ours.” Cloudflare stands by its findings, noting that tests confirmed Perplexity circumvented explicit no-crawl settings. Perplexity has previously faced accusations. Last year, Wired reported the company was plagiarizing news articles, and its CEO struggled to define plagiarism when pressed at a conference.

The Broader Battle Over AI Scraping

As AI services like Perplexity, ChatGPT, and Google Bard rely on vast web data, publishers and site operators have fought back. The robots.txt standard can signal “do not crawl” to search engines and bots. Yet AI companies often ignore these conventions, arguing that public web data is fair use for training models. In May, Cloudflare launched a marketplace letting publishers charge scrapers for access. It also released free tools to help sites block unwanted AI crawlers. CEO Matthew Prince has warned that unregulated scraping threatens the business model of online publishers.

Perplexity
Image Source: Search Engine Journal

What Website Owners Can Do

Cloudflare sites have new options to reject stealth crawlers automatically. Those publishers who have not implemented Cloudflare may use sophisticated bot-management tools or clear-cut server policies in order to implement no-crawl policies. They are able to filter the traffic to look out for signs of the deceitful user-agent spoof, and they are in a position to inhibit requests that resemble the familiar modes of disguise. Incidents in some areas may also take legal actions through terms of service breaches or information protection legislation.

The results provided by Cloudflare point to the conflict between AI innovation and the respect for the preferences of the publishers. With the tactics of data collection maturing on the side of the AI tools, site owners have to remain alert to ensure threats to their content and revenue collection.

TAGGED:AI
Share This Article
Facebook Reddit Copy Link Print
Share
Salman Akhtar
BySalman Akhtar
View enlightening tech pieces written by Salman Keep up with the most recent news, advice, and trends in the field of technology.

Let's Connect

FacebookLike
XFollow
PinterestPin
InstagramFollow
Google NewsFollow
FlipboardFollow

Popular Posts

Microsoft

Microsoft Launches Copilot Tasks Research Preview

Salman Akhtar
4 Min Read
Lenovo

Lenovo Teases Foldable Legion Go With 11.6 Inch Screen

Salman Akhtar
4 Min Read
iOS 26

How iOS 26 Prioritizes Alerts With Apple Intelligence

Salman Akhtar
3 Min Read
Mobile Game Installs

Mobile Game Installs Drop but Spending Climbs in 2025

Salman Akhtar
3 Min Read

You Might Also Like

Nvidia
News

Nvidia posts $68.1B quarter as AI boom accelerates Q4 growth

3 Min Read
Galaxy S26
News

Samsung Reinvents AI on Galaxy S26 With Triple Assistants

4 Min Read
Help Me Calculate Lethal Dose
Blog

How ‘Help Me Calculate Lethal Dose’ Chats Turned Into Cold-Blooded Motel Murders

4 Min Read
Botlash Builds
Blog

Botlash Builds: Neighborhood Campaigns And Lawsuit Armies Hunt Tech For Weakness

4 Min Read

Social Networks

Facebook-f Twitter Instagram Pinterest Rss

Company

  • Home
  • Login
  • My Interests
  • My Saves
  • Register

Policies

  • Home
  • Login
  • My Interests
  • My Saves
  • Register
Latest
Instagram Will Tell Parents When a Teen Searches for Suicide or Self Harm
Gemini Automation Books Rides on Galaxy S26 and Pixel 10
Court Filing Probes Instagram Delay on Teen Protections
YouTube Upgrades Premium Lite With Background Play
Apple Hits New Sales Peak in Europe While Demand Softens Elsewhere

© 2024 The Tech Basic INC. 700 – 2 Park Avenue New York, NY.

TTB White LOGO TB
Follow US
Welcome Back!

Sign in to your account

Username or Email Address
Password

Lost your password?