// Answer

Should I block AI crawlers from my site?

Blocking AI crawlers protects content from training and reuse, but it also keeps you out of AI answers and citations. Here is how to weigh the trade-off and block selectively.

The short answer

Whether you should block AI crawlers depends on your goal, because the same block that protects content from training also keeps you out of AI answers. If visibility in ChatGPT, Perplexity, Gemini, and AI Overviews is a priority, you generally want the answer-facing agents allowed, since an engine cannot cite a page it cannot read. If protecting proprietary or paywalled content is the priority, blocking is reasonable. The practical middle path is selective: block the agents and paths you truly want withheld, allow the ones tied to answers, and decide per agent.

Understand the trade-off

Blocking an AI crawler is a real choice with a real cost. On one side, withholding your content from training and reuse can be the right call for proprietary research, paywalled articles, or material you simply do not want ingested. On the other side, the agents that assemble answers need to fetch your pages to cite them, so a broad block tends to remove you from the answers your customers now see first. There is no universally correct setting, only the one that fits your goal.

Block selectively, not by default

robots.txt supports per-agent and per-path rules, so you rarely need an all-or-nothing block. A common pattern is to allow the answer-facing agents (for example GPTBot, ClaudeBot, PerplexityBot, Google-Extended) on your public, marketing, and reference pages while disallowing sensitive paths like member areas or paywalled content for everyone. Confirm your CDN or firewall enforces the same intent, since an edge rule can block an agent even when robots.txt allows it.

Decide with real data

Before you block broadly, see what you would give up. The free AI visibility scan shows which agents currently reach your site and how it is exposed, and AI citation tracking shows the citations you would lose by closing a door.

See AEO vs traditional SEO for how this fits the wider shift from ranking to being cited.

Frequently asked questions

If I block AI crawlers, will I still show up in AI answers?

Generally no, or much less. Answer engines need to fetch and read a page to quote it, so an agent you block at robots.txt or your firewall usually cannot cite the content behind that block. Blocking is a valid choice when protecting content outranks visibility, but go in knowing the trade-off: the same rule that withholds your pages from reuse also withholds them from citation.

Can I block some AI crawlers and allow others?

Yes. robots.txt lets you set rules per user agent, so you can allow the answer-facing agents (for example GPTBot, ClaudeBot, PerplexityBot, Google-Extended) on your public pages while disallowing others, or block specific paths like paywalled or proprietary sections for everyone. Selective, per-agent and per-path rules are usually smarter than an all-or-nothing block, and you should confirm your CDN or firewall enforces the same intent.

Run your free scan in 60 seconds

Run a free scan to see where you stand across ChatGPT, Claude, Gemini, and Perplexity: which answers cite you, which cite competitors instead, and what to fix first.

Run a free scan