GreyScript AI

AI Crawler robots.txt Generator

Pick which well-known AI crawlers may read your site, add any custom rules, and get a ready-to-paste robots.txt. Training crawlers, on-demand fetch crawlers and search-index crawlers are all labeled, since allowing or blocking each one means something different.

AI crawlers
Each crawler defaults to allowed. Switch a row to Disallow to block it. This list only covers well-known AI crawlers that publish a respected robots.txt user-agent; it is not every bot that might fetch your site, and robots.txt itself is a voluntary convention that well-behaved crawlers choose to follow, not a technical block.
Custom rules (optional)
Add any extra user-agent blocks yourself, for example a crawler not listed above. Pasted as-is above the generated rules.
Sitemap (optional)
Your robots.txt
Nothing you enter is sent anywhere or saved. The file is generated in your browser; copy or download it before you leave the page, then publish it at the root of your domain as /robots.txt.

Questions and answers

What is the difference between a training crawler and an on-demand crawler?

A training crawler like GPTBot or ClaudeBot fetches pages in bulk to build a dataset used to train a model. An on-demand crawler like ChatGPT-User or Claude-User only fetches a specific page when a user of that assistant asks it to read or browse that exact URL. Blocking one does not block the other.

Does robots.txt actually stop an AI crawler from reading my site?

Only if the crawler chooses to follow it. robots.txt is a voluntary convention: well-behaved crawlers from major AI companies generally respect it, but robots.txt cannot technically prevent a request from reaching your server. It is not a security control.

What does Google-Extended actually control?

Google-Extended is not a separate crawler; Googlebot still fetches your pages for Search. Google-Extended is a token that controls whether that already-crawled content may also be used to train Gemini models and power certain AI features, separate from normal Search indexing and ranking.

Why does this tool only list some AI crawlers?

It covers the well-known AI crawlers that publish a respected, documented user-agent string. New crawlers appear and old ones change; use the custom rules box to add any other user-agent your logs show, and check the company's own documentation for the current name before you rely on it.