CIPHERCUE

← AI crawler policy

Which companies block Meta-ExternalAgent in robots.txt?

Meta-ExternalAgent is Meta's AI crawler. CipherCue reads the public robots.txt of tracked organisations and records, per named AI crawler, whether the site declares a rule disallowing it. These are the organisations observed blocking Meta-ExternalAgent. Sign up free, no credit card, to see the full list.

15,812 organisations observed disallowing Meta-ExternalAgent in their robots.txt

A few of the organisations whose robots.txt disallows Meta-ExternalAgent. Sign up free to filter and view the full list.

See all 15,812 organisations blocking Meta-ExternalAgent

Free, no credit card. Filter and view every organisation observed disallowing Meta-ExternalAgent in robots.txt, alongside which other AI crawlers they block and their full observed infrastructure.

How this is observed: CipherCue reads the public robots.txt each organisation publishes and records the disallow rules it declares per named user-agent. A site is recorded as blocking Meta-ExternalAgent when its robots.txt disallows the site root for that agent (or the * catch-all). robots.txt is a stated preference the crawler chooses whether to honour; this records what the site declares, not crawler behaviour. See the AI-crawler methodology.