crawlsignals · How we read · About, corrections and removal · Privacy · records.json
IMDb www.imdb.com · reference
What it says about automated access
robots.txt:
Use of any device, tool, or process designed to data mine or scrape the content using automated means is prohibited without prior written permission from IMDb.
automated-access · source, read 2026-10-03The content on this site is owned by IMDb.com, Inc., an Amazon.com company and is made available for your personal, non-commercial use subject to our Conditions of Use here: https://www.imdb.com/conditions/
reuse, commercial-use · source, read 2026-10-03For authorized data access and licensing inquiries, see: https://www.imdb.com/licensing/
permission, licence · source, read 2026-10-03
| Agent | Rules it follows (RFC 9309) | Over the whole file |
|---|
| * | the * group | disallowed at / (line 69) |
| Googlebot | its own group | 20 Disallow rules; / allowed; not bound by 1 of the * Disallow rules |
| Bingbot | its own group | 20 Disallow rules; / allowed; not bound by 1 of the * Disallow rules |
With no group of their own, these follow the * group: GPTBot, OAI-SearchBot, ChatGPT-User, ClaudeBot, Claude-SearchBot, Claude-User, Google-Extended, Applebot-Extended, PerplexityBot, CCBot, Bytespider, meta-externalagent.
Under RFC 9309 an agent with its own group follows only that group, so the * rules do not apply to it.
Official way in
Not found by us: no API, feed or export.
Documents
Conditions of Use terms · named in robots.txt, line 3 of www.imdb.com · not read by us: robots.txt's comments prohibit crawling, so our crawler reads nothing else from this host
Licensing other · named in robots.txt, line 6 of www.imdb.com · not read by us: robots.txt's comments prohibit crawling, so our crawler reads nothing else from this host
Nothing here is legal advice. A mistake? Write to hello@crawlsignals.com (corrections).