CrawlLedger
Free tool · source-labeled · no input storage

Build an AI crawler policy draft without guesswork.

Choose separate starting rules for documented training, search, archive, and user-requested tokens. CrawlLedger generates ordinary robots.txt groups plus optional Content Signals that you can review before publishing.

Draft, not a one-click legal policy

Replacing an existing robots.txt file can remove search-engine, sitemap, or application rules. Merge carefully, validate the complete file, and check each operator's current documentation.

Crawler access draft

Includes documented training crawlers and non-fetcher controls such as Google-Extended.

Some operators say ordinary robots rules may not apply to a user-triggered fetcher. Omission avoids implying otherwise.

Optional Content Signals

Content Signals express downstream-use preferences. They do not technically block access and are not universally honored.

Exactly which tokens are tracked

The generator and checker use CrawlLedger's versioned registry. A token is tracked only with a first-party operator source and review date.

TokenOperatorPurposeDraft coverage
GPTBotOpenAIModel trainingIncluded when its purpose group is selected
OAI-SearchBotOpenAIAI searchIncluded when its purpose group is selected
OAI-AdsBotOpenAIAd landing-page validationChecker only; configure after reviewing ad requirements
ChatGPT-UserOpenAIUser-requested fetchIncluded when its purpose group is selected
ClaudeBotAnthropicModel trainingIncluded when its purpose group is selected
Claude-SearchBotAnthropicAI searchIncluded when its purpose group is selected
Claude-UserAnthropicUser-requested fetchIncluded when its purpose group is selected
Google-ExtendedGoogleTraining and grounding controlIncluded as a documented control token
Applebot-ExtendedAppleModel trainingIncluded as a documented control token
PerplexityBotPerplexityAI searchIncluded when its purpose group is selected
Perplexity-UserPerplexityUser-requested fetchIncluded when its purpose group is selected
CCBotCommon CrawlOpen web archiveIncluded when its purpose group is selected

OAI-AdsBot is not inserted into generated drafts because OpenAI documents it for advertiser-submitted landing pages, not general search or training. Any robots.txt line is a published instruction for cooperative systems; it does not authenticate a request, enforce behavior, establish consent, or decide legal permission.