Skip to content

robots.txt Generator

Short answer

Create a robots.txt file with blocked paths, sitemaps and a choice for each AI crawler.

Free · No signup · Runs in your browser
Crawling

Start each path with /, e.g. /admin/ or /cart/. Use * as a wildcard and $ to match the end of a URL.

Google ignores Crawl-delay; Bing reads it.

AI crawlers

Crawlers for AI search and answers decide whether your pages can be shown or cited there. Training crawlers collect pages for model training; blocking them does not remove your site from AI search.

AI search and answers
  • OAI-SearchBotOpenAI: indexes pages so its AI search can show and cite them.
  • ChatGPT-UserOpenAI: opens a page when a user asks the assistant about it. Its operator says robots.txt may not apply to these requests.
  • Claude-SearchBotAnthropic: indexes pages so its AI search can show and cite them.
  • Claude-UserAnthropic: opens a page when a user asks the assistant about it.
  • PerplexityBotPerplexity: indexes pages so its AI search can show and cite them.
  • Perplexity-UserPerplexity: opens a page when a user asks the assistant about it. Its operator says robots.txt may not apply to these requests.
Model training
  • GPTBotOpenAI: collects pages to train AI models.
  • ClaudeBotAnthropic: collects pages to train AI models.
  • CCBotOpen web archive crawler; many AI models are trained on its public dataset.
Other AI uses
  • Google-ExtendedGoogle: not a separate crawler — a rule that decides whether pages it already crawls may be used for its AI models.
  • Applebot-ExtendedApple: not a separate crawler — a rule that decides whether pages it already crawls may be used for its AI models.
  • meta-externalagentMeta: crawls pages for its own products; the content may also be used to train AI models.
  • AmazonbotAmazon: crawls pages for its own products; the content may also be used to train AI models.

AI crawler list last checked on 17/09/2026.

Your robots.txt

User-agent: *
Allow: /

What the file allows on the home page

  • GPTBotAllowed
  • OAI-SearchBotAllowed
  • ChatGPT-UserAllowed
  • ClaudeBotAllowed
  • Claude-SearchBotAllowed
  • Claude-UserAllowed
  • PerplexityBotAllowed
  • Perplexity-UserAllowed
  • Google-ExtendedAllowed
  • Applebot-ExtendedAllowed
  • meta-externalagentAllowed
  • AmazonbotAllowed
  • CCBotAllowed

Save the file as robots.txt at the root of your domain (for example https://example.com/robots.txt), then test the live file. Open the robots.txt checker

The file is created in your browser.

What It Does

Creates a robots.txt file with the paths to keep crawlers out of, sitemap lines, an optional Crawl-delay, and an allow or block choice for each known AI crawler, grouped into AI search, model training and other AI uses. It then reads the file back and shows what each AI crawler may fetch.

Why It Matters

robots.txt tells crawlers which URLs they may fetch. AI companies use separate crawlers for search, for fetching a page a user asks about, and for model training, so you can stay visible in AI search while opting out of training. Blocking a training crawler such as GPTBot does not remove a site from ChatGPT search. robots.txt is a request, not access control, and some fetchers that act for a user say they may not follow it.

How It Works

  1. Choose whether to allow crawling or block the whole site

  2. List the paths to keep crawlers out of and your sitemap URLs

  3. Allow or block each AI crawler, or start from a preset

  4. Copy or download the file, save it as /robots.txt and test the live file with the robots.txt checker

Sample input + output

INPUT
crawling: allowed · blocked paths: /admin/, /cart/
AI preset: allow AI search, block the rest
sitemap: https://example.com/sitemap.xml
OUTPUT
User-agent: *
Disallow: /admin/
Disallow: /cart/

User-agent: GPTBot
User-agent: ClaudeBot
User-agent: Google-Extended
User-agent: Applebot-Extended
User-agent: meta-externalagent
User-agent: Amazonbot
User-agent: CCBot
Disallow: /

Sitemap: https://example.com/sitemap.xml

Who Uses This

  • Site owner

    Stay in AI search results while opting out of model training.

  • SEO specialist

    Keep crawlers out of cart, internal search and admin pages without blocking the rest of the site.

  • Web developer

    Create a block-all file for a staging site.

Frequently Asked Questions

Where do I put robots.txt?

At the root of your domain, for example https://example.com/robots.txt. Each subdomain needs its own file.

Does blocking a page in robots.txt remove it from Google?

No. robots.txt controls crawling, not indexing. A blocked URL can still appear in results if other pages link to it; use noindex or password protection to keep a page out.

Which AI crawlers can I control?

The crawler tokens that OpenAI, Anthropic, Perplexity, Google, Apple, Meta and Amazon document, plus CCBot, the crawler of an open web archive. Operators add and retire tokens, so the list is checked regularly; the date is shown under it.

Why don't allowed AI crawlers get their own group?

A crawler with its own group ignores the rules for all crawlers (User-agent: *). Leaving allowed crawlers out keeps your blocked paths in force for them too.

Does Google read Crawl-delay?

No, Google ignores it. Bing reads it.