crawlwise.

Does adobe.com block AI crawlers?

adobe.com serves a bot challenge to 1, and allows 0 (checked 2026-10-01).

The blocks are on research and corpus crawlers rather than on the assistants people use to search, so citation in ChatGPT, Claude and Perplexity is not affected by them.

Every crawler we tested

CrawlerUsed forVerdictWhat we saw
GPTBot
OpenAI
ChatGPT (training)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
OAI-SearchBot
OpenAI
ChatGPT search (retrieval)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
ChatGPT-User
OpenAI
ChatGPT (retrieval)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
ClaudeBot
Anthropic
Claude (training)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
Claude-SearchBot
Anthropic
Claude search (retrieval)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
Claude-User
Anthropic
Claude (retrieval)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
Google-Extended
Google
Gemini (training)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
PerplexityBot
Perplexity
Perplexity (retrieval)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
Perplexity-User
Perplexity
Perplexity (retrieval)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
CCBot
Common Crawl
Common Crawl (research)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
Bytespider
ByteDance
ByteDance (training)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
Applebot-Extended
Apple
Apple Intelligence (training)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
Meta-ExternalAgent
Meta
Meta AI (training)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
Amazonbot
Amazon
Alexa (training)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
cohere-ai
Cohere
Cohere (retrieval)ChallengedChallenged: robots.txt permits it, but the fetch got a bot challenge page instead of the content.
Diffbot
Diffbot
Diffbot (research)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.
Timpibot
Timpi
Timpi (research)Not establishedAllowed by robots.txt; the live fetch timed out, so access was not confirmed.

Homepage SEO snapshot

From the Crawlwise SEO Index run, checked 2026-10-05. One fetch of the homepage, so it describes that page and not the whole site.

Homepagehttps://www.adobe.com/
HTTPSyes
Title60 characters
Meta description133 characters
Canonicalpoints to itself
Noindexno
H1 elements1
Languagenot declared
JSON-LD typesWebSite, Organization, Corporation, ImageObject, PostalAddress
hreflang alternates0
/sitemap.xmlnot found
Response time21 ms

How this was measured

We read adobe.com’s robots.txt, resolved it with an RFC 9309 matcher for each crawler’s product token, then fetched the homepage once as each crawler. A 403 or a challenge page is a block even when robots.txt allows the crawler; a timeout or rate limit is “not established”, never a block. One probe, from one location, on 2026-10-01, so a rule changed since then is not reflected here. Full methodology.

Check your own site

The free checker runs the same probe on any URL: robots.txt resolved for every AI crawler, then a live fetch as each one, so a CDN or firewall block shows up even when robots.txt looks fine.

Check AI crawler access, free

Need it across every page? Site crawls start at $4.99 a month, and Pro adds monitoring that re-checks your URLs on a schedule.

Other sites