crawlwise.

Does ad-score.com block AI crawlers?

ad-score.com blocks all 17 AI crawlers we tested, including GPTBot, ClaudeBot and PerplexityBot (checked 2026-10-01).

Retrieval crawlers are blocked too, so AI assistants that fetch a page live to quote and link it cannot read this site. Its content can still appear in answers from older training data or third-party copies, but not as a fresh, cited source.

Every crawler we tested

CrawlerUsed forVerdictWhat we saw
GPTBot
OpenAI
ChatGPT (training)BlockedBlocked: robots.txt disallows it.
OAI-SearchBot
OpenAI
ChatGPT search (retrieval)BlockedBlocked: robots.txt disallows it.
ChatGPT-User
OpenAI
ChatGPT (retrieval)BlockedBlocked: robots.txt disallows it.
ClaudeBot
Anthropic
Claude (training)BlockedBlocked: robots.txt disallows it.
Claude-SearchBot
Anthropic
Claude search (retrieval)BlockedBlocked: robots.txt disallows it.
Claude-User
Anthropic
Claude (retrieval)BlockedBlocked: robots.txt disallows it.
Google-Extended
Google
Gemini (training)BlockedBlocked: robots.txt disallows it.
PerplexityBot
Perplexity
Perplexity (retrieval)BlockedBlocked: robots.txt disallows it.
Perplexity-User
Perplexity
Perplexity (retrieval)BlockedBlocked: robots.txt disallows it.
CCBot
Common Crawl
Common Crawl (research)BlockedBlocked: robots.txt disallows it.
Bytespider
ByteDance
ByteDance (training)BlockedBlocked: robots.txt disallows it.
Applebot-Extended
Apple
Apple Intelligence (training)BlockedBlocked: robots.txt disallows it.
Meta-ExternalAgent
Meta
Meta AI (training)BlockedBlocked: robots.txt disallows it.
Amazonbot
Amazon
Alexa (training)BlockedBlocked: robots.txt disallows it.
cohere-ai
Cohere
Cohere (retrieval)BlockedBlocked: robots.txt disallows it.
Diffbot
Diffbot
Diffbot (research)BlockedBlocked: robots.txt disallows it.
Timpibot
Timpi
Timpi (research)BlockedBlocked: robots.txt disallows it.

How that compares

How this was measured

We read ad-score.com’s robots.txt, resolved it with an RFC 9309 matcher for each crawler’s product token, then fetched the homepage once as each crawler. A 403 or a challenge page is a block even when robots.txt allows the crawler; a timeout or rate limit is “not established”, never a block. One probe, from one location, on 2026-10-01, so a rule changed since then is not reflected here. Full methodology.

Check your own site

The free checker runs the same probe on any URL: robots.txt resolved for every AI crawler, then a live fetch as each one, so a CDN or firewall block shows up even when robots.txt looks fine.

Check AI crawler access, free

Need it across every page? Site crawls start at $4.99 a month, and Pro adds monitoring that re-checks your URLs on a schedule.

Other sites