crawlsignals · How we read · About, corrections and removal · Privacy · records.json

Facebook www.facebook.com · social

What it says about automated access

robots.txt:

robots.txt https://www.facebook.com/robots.txt read 2026-10-03

AgentRules it follows (RFC 9309)Over the whole file
*the * groupdisallowed at / (line 1352)
Googlebotits own group41 Disallow rules; / allowed; not bound by 1 of the * Disallow rules
Bingbotits own group41 Disallow rules; / allowed; not bound by 1 of the * Disallow rules
GPTBotits own group41 Disallow rules; / allowed; not bound by 1 of the * Disallow rules
ClaudeBotits own group41 Disallow rules; / allowed; not bound by 1 of the * Disallow rules
Google-Extendedits own group41 Disallow rules; / allowed; not bound by 1 of the * Disallow rules
Applebot-Extendedits own group41 Disallow rules; / allowed; not bound by 1 of the * Disallow rules
PerplexityBotits own group41 Disallow rules; / allowed; not bound by 1 of the * Disallow rules

With no group of their own, these follow the * group: OAI-SearchBot, ChatGPT-User, Claude-SearchBot, Claude-User, CCBot, Bytespider, meta-externalagent.

Under RFC 9309 an agent with its own group follows only that group, so the * rules do not apply to it.

Official way in

Not found by us: no API, feed or export.

Documents

Automated Data Collection Terms terms · named in robots.txt, line 7 of www.facebook.com · not read by us: robots.txt's comments prohibit crawling, so our crawler reads nothing else from this host

Nothing here is legal advice. A mistake? Write to hello@crawlsignals.com (corrections).