Technical SEO

Free Robots.txt Checker and Generator

Paste a robots.txt to check whether a URL is allowed for Googlebot, Bingbot and 16 AI crawlers, using the same precedence rules Google follows. Or switch tabs to generate a robots.txt with a clear AI crawler policy.

Free. No signup. Runs in your browser.

Allowed: Googlebot on /blog/pricing-guide

No rule matches this path, so crawling is allowed.

AI crawlerUsed forThis path
GPTBot · OpenAIModel trainingBlocked
OAI-SearchBot · OpenAIAI search indexAllowed
ChatGPT-User · OpenAILive user requestAllowed
ClaudeBot · AnthropicModel trainingAllowed
Claude-SearchBot · AnthropicAI search indexAllowed
Claude-User · AnthropicLive user requestAllowed
PerplexityBot · PerplexityAI search indexAllowed
Perplexity-User · PerplexityLive user requestAllowed
Google-Extended · GoogleModel trainingAllowed
Applebot-Extended · AppleModel trainingAllowed
meta-externalagent · MetaModel trainingAllowed
Amazonbot · AmazonAI search indexAllowed
Bytespider · ByteDanceModel trainingAllowed
CCBot · Common CrawlModel trainingBlocked
DuckAssistBot · DuckDuckGoLive user requestAllowed
MistralAI-User · MistralLive user requestAllowed

What you get

Allowed or blocked, with the reason

The exact rule and group that decided it, using RFC 9309 matching.

AI crawler access table

GPTBot, ClaudeBot, PerplexityBot, Google-Extended and more, checked in one view.

Generator with presets

Allow AI search while blocking AI training in one click, then copy the file.

How robots.txt rules are matched

A crawler follows only the group whose User-agent line best matches its name, and falls back to the User-agent: * group only when no specific group exists. That surprises people: if you add a GPTBot group, GPTBot ignores everything in your * group.

Inside the chosen group, the longest matching path wins. If an Allow and a Disallow rule match with the same length, Allow wins. An asterisk matches any characters and a dollar sign anchors the end of the URL. This checker applies those rules exactly, so what you see is what Googlebot and well-behaved AI crawlers do.

AI training vs AI search crawlers

AI companies now run separate crawlers for separate jobs. GPTBot, ClaudeBot and Google-Extended collect content for model training. OAI-SearchBot, Claude-SearchBot and PerplexityBot build the indexes that AI answers cite. ChatGPT-User and Claude-User fetch a page live when a person asks about it.

Blocking all of them keeps you out of AI answers entirely. Many sites choose the middle path: block training crawlers, allow search and user-triggered ones so the brand still shows up and gets cited. The generator's default preset does exactly that.

Common robots.txt mistakes

The most expensive one is a staging file shipped to production with Disallow: / under User-agent: *, which removes the whole site from search. Others: blocking CSS and JavaScript that Google needs to render pages, using robots.txt to hide pages that are linked elsewhere (use noindex instead), and forgetting the Sitemap line.

Frequently asked questions

How do I check if my robots.txt blocks a page?

Paste your robots.txt, enter the URL path and pick a crawler. The checker shows allowed or blocked and the exact rule that decided it.

Should I block GPTBot?

Block GPTBot if you do not want your content used for OpenAI model training. Blocking it does not remove you from ChatGPT search, which uses OAI-SearchBot and ChatGPT-User.

Does robots.txt remove a page from Google?

No. It stops crawling, not indexing. A blocked page can still appear in results if other sites link to it. Use a noindex tag to keep a page out of the index.

Does Google support crawl-delay?

No. Googlebot ignores crawl-delay. Some other crawlers respect it.

RunAgents

Get alerted the moment robots.txt changes

RunAgents watches your robots.txt and AI bot access and alerts you when a deploy blocks Google or an AI crawler you depend on.

Get started →

More free tools

See all free marketing tools →