crawlsignals · How we read · About, corrections and removal · Privacy · records.json

BBC www.bbc.com · publisher

What it says about automated access

robots.txt:

robots.txt https://www.bbc.com/robots.txt read 2026-10-03

AgentRules it follows (RFC 9309)Over the whole file
*the * group49 Disallow rules; / allowed
GPTBotits own groupdisallowed at / (line 154), with 2 Allow rules; not bound by 1 of the * Disallow rules
OAI-SearchBotits own groupdisallowed at / (line 181)
ChatGPT-Userits own groupdisallowed at / (line 159), with 2 Allow rules; not bound by 1 of the * Disallow rules
ClaudeBotits own groupdisallowed at / (line 133)
Google-Extendedits own groupdisallowed at / (line 164), with 2 Allow rules; not bound by 1 of the * Disallow rules
Applebot-Extendedits own groupdisallowed at / (line 151)
PerplexityBotits own groupdisallowed at / (line 169)
CCBotits own groupdisallowed at / (line 121)
Bytespiderits own groupdisallowed at / (line 142)
meta-externalagentits own groupdisallowed at / (line 178)

With no group of their own, these follow the * group: Googlebot, Bingbot, Claude-SearchBot, Claude-User.

Under RFC 9309 an agent with its own group follows only that group, so the * rules do not apply to it.

Official way in

Not found by us: no API, feed or export.

Documents

Terms of Use terms · named in robots.txt, line 3 of www.bbc.co.uk · not read by us: the robots.txt of www.bbc.co.uk prohibits crawling in its comments, so our crawler reads nothing else there

Can I use BBC content for my business? policy · named in robots.txt, line 15 of www.bbc.co.uk · not read by us: the robots.txt of www.bbc.co.uk prohibits crawling in its comments, so our crawler reads nothing else there

Nothing here is legal advice. A mistake? Write to hello@crawlsignals.com (corrections).