Skip to content

generator · free · no sign-up

robots.txt generator

Write crawl rules for search and retrieval bots. Training preferences are a separate decision.

Answers: robots.txt generator · block gptbot robots.txt · disallow all robots

Read the step-by-step guide · Blog workflow

What this does not measure

  • robots.txt is a request, not enforcement. A crawler that ignores it is not stopped by it.
  • It does not know which paths exist on your site. Ask the crawl for that.
  • Blocking a training crawler and blocking a search crawler are different decisions with different consequences, and the generator does not make either one for you.
Examples are fictional templates. Replace their details before using the output.
Rule 1

One per line. Use * for the catch-all group.

Absolute URLs, one per line.

  • No Sitemap declared. Add the absolute URL of your sitemap.xml.
robots.txt
User-agent: *

Crawler reference

User-agentOwnerCategory
GooglebotGooglesearch
Googlebot-ImageGooglesearch
Google-ExtendedGoogletraining
AdsBot-GoogleGooglesearch
BingbotMicrosoftsearch
MicrosoftPreviewMicrosoftretrieval
DuckDuckBotDuckDuckGosearch
ApplebotApplesearch
Applebot-ExtendedAppletraining
GPTBotOpenAItraining
OAI-SearchBotOpenAIretrieval
ChatGPT-UserOpenAIretrieval
ClaudeBotAnthropictraining
Claude-SearchBotAnthropicretrieval
Claude-UserAnthropicretrieval
PerplexityBotPerplexityretrieval
Perplexity-UserPerplexityretrieval
Google-CloudVertexBotGoogletraining
Meta-ExternalAgentMetatraining
Meta-ExternalFetcherMetaretrieval
facebookexternalhitMetasocial
BytespiderByteDancetraining
CCBotCommon Crawlresearch
DiffbotDiffbotresearch
Cohere-aiCohereretrieval
cohere-training-data-crawlerCoheretraining
MistralMistraltraining
GrokxAIretrieval
AmazonbotAmazontraining
YandexBotYandexsearch
BaiduspiderBaidusearch
TwitterbotXsocial
LinkedInBotLinkedInsocial
SlackbotSlacksocial
DiscordbotDiscordsocial
WhatsAppMetasocial
AhrefsBotAhrefsseo
SemrushBotSemrushseo
DotBotMozseo