robots.txt AI generator
Decide which AI crawlers can read your site. Toggle
GPTBot, ClaudeBot,
PerplexityBot, Google-Extended and more,
add your sitemap, then copy or download a ready-to-ship
robots.txt. Runs entirely in your browser — no signup.
How it works
- Choose a policy. Start from a preset or toggle each
crawler. Allowed emits
Allow: /; Blocked emitsDisallow: /for that user-agent. - Add your sitemap (optional) so crawlers can discover
your pages — one
Sitemap:line is added to the file. - Ship it. Copy or download, then add the lines to
robots.txtat your domain root. If you already have one, append these blocks rather than replacing the file.
Which crawlers should I allow?
If your goal is to get cited by AI answer engines,
allow the answer-engine crawlers (GPTBot, PerplexityBot, ClaudeBot,
Google-Extended and friends) — that's how ChatGPT and Perplexity learn
your pages exist. If your content is your product and you don't want it
used for bulk model training, block dataset crawlers
like CCBot and Bytespider. The
Recommended preset does exactly this split.
Allowing a crawler in robots.txt is necessary but not
sufficient to actually get cited — answer engines also need to
understand your pages. Run the free
AICite audit to see all 13 signals, or build your
llms.txt with the free generator.
FAQ
Is this generator free?
Yes — completely free, no signup, and it runs entirely in your
browser. The $24 Pro Report bundles the
robots.txt patch with llms.txt,
llms-full.txt, JSON-LD, and an AI-Ready badge.
Should I allow or block AI crawlers?
Allow answer engines if you want AI citations; block bulk trainers if your content is a paid product. The worst option is no rule at all — an implicit allow you never chose. The Recommended preset is a safe default for most sites that want traffic and citations.
Do I replace my existing robots.txt?
No — append these User-agent blocks to your current
robots.txt. Keep any existing
User-agent: * group and your existing
Sitemap: line if you already have one.
How do I verify it's working?
Fetch it with a spoofed agent:
curl -A "GPTBot" -I https://yourdomain.com/robots.txt.
Then run the free AICite audit and check the
/robots.txt signal.
Built by AICite — the free A–F audit for how AI answer engines see your site. Configure your policy above, then grade your whole site free.