Skip to content
Free robots.txt generator

Build a robots.txt file and test it.

Choose what search engines and AI crawlers may read, then copy or download the file and test any path against it. Free, no account, unlimited.

One per line. Ignored when all crawlers are blocked.

AI crawlers

Model training

Blocking these keeps your pages out of future training data. It does not remove you from AI answers.

  • GPTBot OpenAI

    Collects pages to train OpenAI's models.

  • ClaudeBot Anthropic

    Collects pages to train Anthropic's Claude models.

  • Google-Extended Google

    Decides whether Google may train Gemini on your pages. Google Search is not affected.

  • Applebot-Extended Apple

    Decides whether Apple may train its models on your pages. Siri and Spotlight are not affected.

  • CCBot Common Crawl

    Builds the open Common Crawl dataset many AI models train on.

  • Bytespider ByteDance

    Collects pages to train ByteDance's models. Reported not to always follow robots.txt.

  • meta-externalagent Meta

    Collects pages to train Meta's AI models and improve products.

  • Diffbot Diffbot

    Turns pages into structured data sold to AI and analytics teams.

AI search and assistants

These find and read pages for answers in ChatGPT, Claude, Perplexity and others. Block them and those answers cannot cite you.

  • OAI-SearchBot OpenAI

    Indexes pages so ChatGPT search can show and link to them.

  • ChatGPT-User OpenAI

    Opens a page when a ChatGPT user asks about it.

  • Claude-SearchBot Anthropic

    Indexes pages so Claude can find and cite them in answers.

  • Claude-User Anthropic

    Opens a page when a Claude user's question needs it.

  • PerplexityBot Perplexity

    Indexes pages so Perplexity can show and cite them.

  • Perplexity-User Perplexity

    Opens a page when a Perplexity user asks. Perplexity says these visits may ignore robots.txt.

  • Meta-ExternalFetcher Meta

    Opens a link when a Meta AI user asks about it.

  • Amazonbot Amazon

    Reads pages for Amazon services, including answers in Alexa.

  • DuckAssistBot DuckDuckGo

    Reads pages for DuckDuckGo's AI assisted answers.

  • MistralAI-User Mistral

    Opens a page when a Le Chat user asks about it.

Free, no account, unlimited. The file is built in your browser and nothing is sent anywhere.

  • Rules for all crawlers
  • 18 AI crawlers, one switch each
  • Training and AI search kept apart
  • Copy or download the file
  • Built in tester for any path

What you get

A file you can upload, and proof it does what you meant.

The generator writes a clean robots.txt from your choices, sums up what it allows and blocks, and lets you test paths against it before it goes live.

robots.txt · example.com

ready

  • all crawlersallowed
  • disallowed paths/admin/, /cart
  • training crawlers8 of 8 blocked
  • AI answer agents10 of 10 allowed, still cited
  • sitemaplisted
  • test /admin/ as Googlebotblocked by line 2

Why it matters

One small file decides what gets read.

Search engines and AI crawlers check robots.txt before they request a page. A good file points them at your best pages. A bad one can hide your whole site.

  1. Crawling

    Keep crawlers out of clutter

    Admin screens, carts and internal search results waste crawl time. A disallow rule keeps crawlers on the pages you want found.

  2. AI visibility

    Choose what AI may learn and cite

    Training crawlers and AI search agents use separate user agents. You can opt out of training and still be read, cited and linked in AI answers.

  3. Safety

    Avoid blocking your whole site

    One wrong line, such as Disallow: / for every user agent, hides every page from search. Testing paths before you upload catches it.

Who it's for

Built for the people who run websites.

01

Site owners

Get a correct robots.txt without learning the syntax, and decide in one place how AI tools may use your content.

02

SEO specialists

Draft rules for a client site, test the paths that matter, and hand over a file that will not block pages by accident.

03

Developers

Generate the file for a new build or a staging site, download it, and drop it in the public root.

04

Publishers

Opt out of AI training while staying visible in ChatGPT search, Claude and Perplexity answers that send readers back.

How it works

From choices to a tested file in a minute.

  1. 01

    Set the basics

    Pick allow or block for all crawlers, list the paths to keep private, and add your sitemap URL and a crawl delay if you want them.

  2. 02

    Choose your AI crawlers

    Allow or block each one, or use a shortcut. Training crawlers and AI answer agents are listed apart, so the tradeoff stays clear.

  3. 03

    Copy, test and upload

    Copy or download robots.txt, test any path against it as any crawler, then upload it to the root of your site.

FAQ

Questions, answered.

Free SEO tools

See where your redirects really lead next.