CIPHERCUE

← AI crawler policy

Which companies block PerplexityBot in robots.txt?

PerplexityBot is Perplexity's AI crawler. CipherCue reads the public robots.txt of tracked organisations and records, per named AI crawler, whether the site declares a rule disallowing it. These are the organisations observed blocking PerplexityBot. Sign up free, no credit card, to see the full list.

14,700 organisations observed disallowing PerplexityBot in their robots.txt

A few of the organisations whose robots.txt disallows PerplexityBot. Sign up free to filter and view the full list.

See all 14,700 organisations blocking PerplexityBot

Free, no credit card. Filter and view every organisation observed disallowing PerplexityBot in robots.txt, alongside which other AI crawlers they block and their full observed infrastructure.

How this is observed: CipherCue reads the public robots.txt each organisation publishes and records the disallow rules it declares per named user-agent. A site is recorded as blocking PerplexityBot when its robots.txt disallows the site root for that agent (or the * catch-all). robots.txt is a stated preference the crawler chooses whether to honour; this records what the site declares, not crawler behaviour. See the AI-crawler methodology.