BeamciteSaaS

Is ChatGPT even allowed to read your site?

Enter your domain and we check whether the crawlers behind ChatGPT, Claude, Perplexity and Gemini can actually read it, plus whether you serve an llms.txt. If AI crawlers are blocked, no amount of good content gets you cited. Free, instant, no sign-up.

Try:

01 / How it works

Step 1

Enter your domain

Just the domain, no need for https:// or a path. We fetch the live robots.txt from the site root.

Step 2

We check 12 AI crawlers

Every request is matched against robots.txt using the same group-selection and path-matching logic the crawlers themselves use, so the verdict is accurate.

Step 3

See exactly what's blocking you

Each bot gets an allowed, partial or blocked verdict, with the exact rule line when one applies, plus a check for llms.txt.

02 / FAQ

Questions, answered.

What is an AI crawler, and why would it be blocked?+

AI crawlers are the bots that feed AI answer engines: GPTBot and OAI-SearchBot for ChatGPT, ClaudeBot for Claude, PerplexityBot for Perplexity, Google-Extended for Gemini, and others. Sites sometimes block them on purpose (to opt out of AI training) or by accident, often through a blanket 'Disallow: /' rule meant for something else entirely.

If a crawler is blocked, does that mean I can never be cited by that AI?+

In almost all cases, yes for that specific crawler. If GPTBot cannot read your page, ChatGPT has no way to know the page exists, let alone quote it. This is the first, most basic gate: being crawlable does not guarantee a citation, but being blocked guarantees you cannot get one.

What is the difference between a bot like GPTBot and ChatGPT-User?+

GPTBot and similar bots crawl your site ahead of time to build a training or search index. Bots named with '-User' (ChatGPT-User, Perplexity-User, Claude-User) instead fetch a page live, in the moment someone asks that assistant a question about it. Blocking one but not the other has very different effects, which is why we check all 12 separately.

Does Google-Extended affect my Google Search rankings?+

No. Google-Extended only controls whether your content can be used to train and ground Google's Gemini models and AI Overviews; it has no effect on classic Google Search ranking. Beamcite does not sell or promise Google rankings; this checker (and Beamcite) is strictly about visibility in AI assistants.

What is llms.txt, and do I need one?+

llms.txt is a plain-text file at your domain root that maps your site for AI crawlers: what it is, and which pages matter most. It is an emerging convention, not a requirement, and no assistant is obligated to read or honor it. It can make your site easier for supporting crawlers to understand quickly, nothing more.

I found a blocked crawler. How do I fix it?+

Edit your robots.txt (usually at the site root, e.g. yoursite.com/robots.txt) and remove or narrow the Disallow rule under that bot's User-agent group. If you cannot find a dedicated group for it, it is likely inheriting a blanket rule under 'User-agent: *'.

Does this tool store the domains people check?+

No. We fetch the domain's robots.txt live, server-side (browsers cannot fetch other sites directly), run the checks, and return the result. Nothing is logged against your domain or saved.

Free strategy call · 30 minutes

Crawlable is step one. Cited is the goal.

Beamcite makes sure AI crawlers can read you, then does the ongoing work of getting you actually cited by ChatGPT, Perplexity and Gemini, with a human running it.

Book the call

No payment. No obligation. A real human.