How Do I Block AI Crawlers From My Site?
Blocking AI crawlers means adding disallow rules to a site's robots.txt file, naming each bot specifically. Common ones include GPTBot and OAI-SearchBot from OpenAI, ClaudeBot from Anthropic, PerplexityBot, and Google-Extended, which covers Gemini and AI Overviews training separately from regular Googlebot.
Blocking these bots keeps a site out of that tool's training data or live citations, but it is an all-or-nothing switch per crawler. There is no way to allow citation while blocking training for the same bot on most platforms.
Most sites chasing AI visibility do the opposite: they check that these crawlers are explicitly allowed, since a leftover blanket disallow rule from years ago can quietly shut a site out of every AI answer engine.
Checking crawler access is one of the first things LLM optimization software verifies on a new site.
BlazeHive finds the buyer-intent keywords your customers search right before they buy, then writes and publishes the pages that put you in front of them.
Start free trial →