Free Chrome extension

Check if AI crawlers can access your website

Check the robots.txt rules affecting ChatGPT, Claude, Perplexity and Google AI-related crawler controls directly from your browser.

Chrome Web Store listing coming soonExtension privacy

Free · No account · No tracking

Check AI crawler access from any page

Most crawler checks only look at a domain. The checker reads the robots.txt published at the exact origin you are on — scheme, host and port — then evaluates the rules against the page path you are actually viewing, including its query string. A site can allow crawling at the root and still block a template, a section or a single document, so a domain-level answer is often wrong for the page that matters.

Matching follows the Robots Exclusion Protocol: specific user-agent groups take precedence over the wildcard group, * and $ patterns are honoured, and the most specific matching rule wins, with Allow preferred when rules are equally specific.

Supported crawler controls

OAI-SearchBot
ChatGPT Search

OpenAI's crawler for ChatGPT search results. Controlling this token is separate from controlling GPTBot.

GPTBot
OpenAI model development

OpenAI's crawler used for model development. Blocking GPTBot does not block ChatGPT Search.

Claude-SearchBot
Claude Search

Anthropic's crawler for Claude search results. Separate from ClaudeBot.

ClaudeBot
Anthropic model development

Anthropic's crawler used for model development. Blocking ClaudeBot does not block Claude Search.

Claude-User
Claude user-requested retrieval

Used when a Claude user asks for a specific page to be fetched during a conversation.

PerplexityBot
Perplexity Search

Perplexity's crawler for search and discovery.

Googlebot
Google Search

Google's primary search crawler.

Google-Extended
Gemini use control

A robots.txt control token, not a crawler. It controls whether crawled content may be used for Gemini and related Google AI products. It does not control Google Search crawling.

What the checker can and cannot tell you

It can tell you

  • Whether a crawler token is allowed or blocked by robots.txt for the current URL.
  • Which rule and which user-agent group produced that outcome.
  • Whether a robots.txt file exists at the origin at all.
  • When the answer cannot be verified, rather than guessing.

It cannot tell you

  • Whether a crawler can actually reach your site. A WAF, CDN, IP reputation rule or bot protection service can block a crawler that robots.txt permits.
  • Whether a crawler chooses to crawl the page, or how often. Permission is not a schedule.
  • Whether an AI engine mentions, recommends or cites your brand. Crawler permission is not visibility.

Results use four precise statuses: Allowed by robots.txt, Blocked by robots.txt, No robots.txt restriction detected and Unable to verify. There is no score and no percentage, because robots.txt does not support one.

Why crawler access matters

Crawler access is a technical prerequisite. If a search or model crawler is blocked from a page, that page cannot be retrieved through that route at all — so any downstream outcome is off the table before content, structure or authority is considered. Blocking is often unintentional: rules copied between environments, a staging directive shipped to production, or a broad Disallow written before AI crawler tokens existed.

Checking access tells you the door is open. It does not tell you anyone walked through it.

Want to understand whether AI engines actually mention, recommend and cite your brand?

Citations.io tracks how ChatGPT, Claude, Gemini and Perplexity answer the prompts your buyers ask, and what to change each month.