← Back to the library
Glossary AI CrawlersAI AnswersTechnical

AI Crawlers Explained: GPTBot, ClaudeBot, PerplexityBot & Google-Extended

AI crawlers are the bots that read your website for AI engines. Block them and you can't be named. Here's what each one is and how to let them in — in plain English.

The short answer

AI crawlers are the automated bots — like GPTBot, ClaudeBot, PerplexityBot, and Google-Extended — that read your website so AI engines can understand and cite it. Blocking them in your site's robots.txt file is the one technical mistake that makes it impossible for AI to name you, and it's easy to check and fix.

AI crawlers are the automated bots that read your website so AI engines can understand it and cite it in their answers. If you block them, you can’t be named — it’s that direct.

Crawler access is the one genuinely technical gate in all of this. Most of getting named by AI is about clarity and trust, not code. But this one is a simple yes/no: can the bots read your site or not?

Not sure whether you’re letting the AI crawlers in? Run a free AI Readiness Glimpse — it checks crawler access as part of your site’s signals.

What is an AI crawler?

A crawler is a program that visits web pages and reads their content. Search engines have used crawlers for decades. AI companies now run their own so their models can find and understand current information — including who the best local business is for a given job.

When you allow these crawlers, you’re giving the AI a clean copy of your facts to work from. When you block them, the AI is working with nothing, or with stale second-hand information.

The AI crawlers worth knowing

You don’t need to memorize these, but here are the main ones and who they belong to:

  • GPTBot — OpenAI’s crawler, which feeds ChatGPT.
  • ClaudeBot — Anthropic’s crawler, which feeds Claude.
  • PerplexityBot — Perplexity’s crawler, which feeds its answer engine.
  • Google-Extended — Google’s setting that controls whether your content can be used for its AI features.

There are others, but these are the ones that matter most for a business that wants to show up when buyers ask AI for a recommendation.

How to check — and fix — crawler access

Your site has a small file called robots.txt that tells crawlers what they’re allowed to read. You can see yours by visiting yourdomain.com/robots.txt in a browser.

  • If it says something like Disallow: / under one of those bot names, that bot is being turned away.
  • To let AI in, that file should allow the crawlers above rather than block them.

This often gets set the wrong way by accident — a privacy plugin, a site template, or a setting someone flipped years ago. It’s a two-minute fix once you know it’s there, which is exactly why we check it for you.

The honest version

Letting the crawlers in doesn’t guarantee AI will name you — it just makes it possible. Think of it as unlocking the door. Everything else on our site is about giving AI a reason to walk through it and pick you: clear facts, direct answers, and real proof. But if the door’s locked, none of the rest matters.

Keep reading: Why isn’t my business showing up in AI answers? · What is AI visibility? · Does llms.txt work?

Free

Get My Free Glimpse Report

See whether your visible site signals have obvious AI-visibility gaps before you invest in the AI True Score Full Report.

Get My Free Glimpse Report →

Full report

Unlock the AI True Score Full Report

Go deeper than the Glimpse with 59 AI-visibility signals and 10 buyer-intent prompts across ChatGPT, Gemini, Perplexity, Claude, and Grok model families.

Unlock the Full Report →

Frequently asked questions

What happens if I block AI crawlers?

If you block the crawlers, the AI engines behind them can't read your site — so they have nothing to understand, trust, or name. Blocking is one of the few things that guarantees invisibility in AI answers, and it often happens by accident through a plugin, a template, or a well-meaning setting.

How do I know if I'm blocking AI crawlers?

Check the robots.txt file at yourdomain.com/robots.txt and look for lines that disallow bots like GPTBot, ClaudeBot, PerplexityBot, or Google-Extended. If you're not sure what you're looking at, a free AI Readiness Glimpse checks crawler access for you as part of your site's readiness signals.

Should I allow all the AI crawlers?

For most local and service businesses that want to be recommended by AI, yes — allowing the major AI crawlers is what lets those engines read and cite you. Blocking them only makes sense if you have a specific reason to keep your content out of AI systems, which is rare for a business that wants customers.

Does allowing crawlers guarantee I'll be recommended?

No. Allowing crawlers is necessary but not sufficient — it opens the door, but AI still has to understand, trust, and choose you over others. Blocking them, though, guarantees you're left out. Access first, everything else second.