CrawlLedger
Free tool · source-labeled · no input storage

Build an AI crawler policy draft without guesswork.

Choose separate starting rules for documented training, search, archive, and user-requested tokens. CrawlLedger generates ordinary robots.txt groups plus optional Content Signals that you can review before publishing.

Draft, not a one-click legal policy

Replacing an existing robots.txt file can remove search-engine, sitemap, or application rules. Merge carefully, validate the complete file, and check each operator's current documentation.

Crawler access draft

Includes documented training crawlers and non-fetcher controls such as Google-Extended.

Some operators say ordinary robots rules may not apply to a user-triggered fetcher. Omission avoids implying otherwise.

Optional Content Signals

Content Signals express downstream-use preferences. They do not technically block access and are not universally honored.

Exactly which tokens are included

The generator is built from CrawlLedger's versioned registry. A token is included only with a first-party operator source and review date.

TokenOperatorPurposeFetcher or control
GPTBotOpenAIModel trainingDocumented fetcher
OAI-SearchBotOpenAIAI searchDocumented fetcher
ChatGPT-UserOpenAIUser-requested fetchDocumented fetcher
ClaudeBotAnthropicModel trainingDocumented fetcher
Claude-SearchBotAnthropicAI searchDocumented fetcher
Claude-UserAnthropicUser-requested fetchDocumented fetcher
Google-ExtendedGoogleTraining and grounding controlControl token; not a separate fetcher
Applebot-ExtendedAppleModel trainingControl token; not a separate fetcher
PerplexityBotPerplexityAI searchDocumented fetcher
Perplexity-UserPerplexityUser-requested fetchDocumented fetcher
CCBotCommon CrawlOpen web archiveDocumented fetcher

A robots.txt line is a published instruction for cooperative systems. It does not authenticate a request, enforce behavior, establish consent, or decide legal permission.