AI Crawler Access Auditor
Audit whether AI bots (GPTBot, ClaudeBot, PerplexityBot, Google-Extended +more) can crawl your site, and generate the correct robots.txt + llms.txt. Free, no API key.
AI Crawler Access Checker
Enter your URL and see whether eight named AI crawlers are allowed or blocked in your robots.txt. You get a 0-100 score, a count of how many AI search engines can reach you, and a ready-to-paste robots.txt block plus a starter llms.txt. Free, no signup.
How do I check if AI crawlers can access my website?
Open your robots.txt file and look for rules naming AI crawlers such as GPTBot, OAI-SearchBot, ClaudeBot, Claude-SearchBot and PerplexityBot. A Disallow rule under any of those names blocks that bot. This free checker reads the file for you, gives a verdict for each of eight crawlers, and generates a corrected robots.txt block you can paste in.
Training crawlers and search crawlers are not the same thing
This is the single most important idea on this page, and it is easy to get wrong. Some AI bots collect text to train models. Others fetch pages so they can cite you in an answer. Blocking the first group is a fair business decision. Blocking the second group takes you out of AI answers completely, and that can happen by accident with one blanket rule.
Search crawlers: let these in
OAI-SearchBot, Claude-SearchBot, PerplexityBot and ChatGPT-User fetch pages so an AI can link to you. OpenAI says OAI-SearchBot does not collect training data. If you block these, you are not protecting your content, you are removing yourself from the answers your customers see.
Training crawlers: your call
GPTBot, ClaudeBot, Google-Extended and Applebot-Extended govern whether your text is used to train models. Publishers with original research often block them. For a local plumber or dentist, there is usually little to lose and some upside in leaving them open. The checker reports them separately so you decide.
Google-Extended is not Googlebot
Google-Extended is a robots.txt token that controls whether your content trains Gemini. Google states it does not affect your inclusion in Google Search, and AI Overviews are built from Googlebot's crawl. So blocking Google-Extended does not keep you out of AI Overviews, and it does not hurt your rankings.
The blanket-block trap
A rule reading User-agent: * followed by Disallow: / blocks everything, including the search crawlers that would have cited you. Plugins and security tools sometimes add AI-blocking rules during setup without making the training and search difference clear. The checker shows the verdict per bot so nothing hides.
The eight crawlers we check
Each line is the exact user-agent name to use in your robots.txt file.
A score, a verdict list, and files you can paste
The output is meant to be usable in five minutes, not studied. You get a 0-100 access score, a yes or no for each of the eight crawlers, a plain count of how many AI search engines can currently reach you, and two generated files. Copy the robots.txt block over your current AI rules and you are done.
The generated robots.txt block
We build a block that names each AI crawler explicitly, allows the search bots, and leaves the training bots as a choice you make. Paste it into your existing robots.txt rather than replacing the whole file, so your other rules and sitemap line survive.
The starter llms.txt
We also generate a simple llms.txt listing your key pages. Be realistic about it: Google has said it does not use llms.txt for Search or AI Overviews, and no major AI company has formally committed to reading it in production. It is cheap insurance, not a ranking lever.
Where to put the file
Robots.txt has to sit at the root of your domain, so it loads at yoursite.com/robots.txt. On WordPress many SEO plugins have a robots.txt editor built in. On other platforms you upload it to the public folder. If it loads anywhere else, crawlers will not find it.
AI Crawler Access Checker FAQ
Should I block AI crawlers from my website?
What is the difference between GPTBot and OAI-SearchBot?
Will blocking Google-Extended hurt my Google rankings?
Why does Applebot-Extended never show in my logs?
I do not have a robots.txt file at all. Is that bad?
Do AI crawlers actually obey robots.txt?
How often should I re-run this check?
Other free tools worth a minute
Not sure what your robots.txt should say?
Run the check, then send us the result. On a free strategy call we will tell you which crawlers to allow for your business and why, without selling you a service you do not need.
Book a free CEO strategy call →