How it works

How to check which AI crawlers you allow

  1. 01

    Enter your website

    Type a domain or a full page address. We read the robots.txt at the root of that site.

  2. 02

    Rules are matched per crawler

    Each AI crawler gets its own verdict for that exact page, using the same matching rules the crawlers follow (RFC 9309).

  3. 03

    Fix what blocks you

    If an AI search crawler is shut out, copy the lines that let it in. You can also paste edited rules and test them before you deploy.

The crawlers we check

Not every AI bot does the same job

Blocking a search crawler hides you from that assistant's answers. Blocking a training crawler only keeps your pages out of model training. Most robots.txt mistakes come from mixing the two up.

User agentJob
OAI-SearchBotAI search
Claude-SearchBotAI search
PerplexityBotAI search
ChatGPT-UserOn request
GPTBotTraining
ClaudeBotTraining
Google-ExtendedUsage control
Applebot-ExtendedUsage control
Meta-ExternalAgentTraining
CCBotTraining
BytespiderTraining

Learn more

After the door is open

FAQ

AI crawler and robots.txt questions

Is this AI crawler checker free?

Yes. There is no signup and no email. Enter a website and you get the verdict for every AI crawler we track in a few seconds.

How do I check if my website blocks ChatGPT?

Enter your site above and look at three rows: OAI-SearchBot, which indexes pages for ChatGPT search, ChatGPT-User, which opens a page when someone asks ChatGPT to read it, and GPTBot, which collects training data. If OAI-SearchBot is blocked, ChatGPT search cannot index your pages.

What is the difference between GPTBot and OAI-SearchBot?

GPTBot collects public pages that may be used to train OpenAI models. OAI-SearchBot indexes pages so ChatGPT can show and link them in search answers. They are controlled separately, so you can block GPTBot to opt out of training and still allow OAI-SearchBot to stay visible in ChatGPT search.

Should I block AI crawlers?

It depends on what you want. Blocking training crawlers such as GPTBot, ClaudeBot and CCBot is a choice about how your content is used. Blocking search crawlers such as OAI-SearchBot, Claude-SearchBot and PerplexityBot keeps you out of those assistants' answers. Many sites allow the search crawlers and decide on training separately.

Does blocking Google-Extended hurt my Google rankings?

No. Google-Extended is a robots.txt token, not a separate crawler. Google states it does not affect inclusion or ranking in Google Search. It controls whether your content is used for Gemini training and grounding. AI Overviews in Google Search are crawled by Googlebot, not Google-Extended.

My robots.txt allows AI crawlers. Can they still be blocked?

Yes. robots.txt is only one layer. A CDN or firewall rule, a bot-protection setting or a login wall can turn crawlers away even when robots.txt says yes. This tool reads robots.txt only. The free AI visibility scan requests the page itself the way a crawler does and shows what actually comes back.

What does Allowed by default mean?

No rule in your robots.txt mentions that crawler or matches that page, so the crawler may visit. That is the normal state for most sites and nothing needs fixing.

What is llms.txt and do I need one?

llms.txt is a proposed plain-text file at the root of a site that gives AI tools a short, curated map of your most useful pages. It is not an official standard and support from AI companies is limited so far, so treat it as optional. We show whether one exists so you know where you stand.

One file is not the verdict.

robots.txt decides whether AI may visit. The free AI visibility scan shows what it finds when it does, graded A to F with the exact fixes.

Check another siteRun the free AI scan