Should I block AI crawlers from my website?

Blocking AI crawlers usually hides you from AI answers without erasing what models know. Check yours free with Prefer's Crawler Access Checker.

Javed Khatri Javed Khatri Co-founder, Prefer

2 min read Answer

The short answer

Should I block GPTBot in my robots.txt?

Usually no. Blocking search crawlers like OAI-SearchBot or PerplexityBot removes you from the AI answers buyers read, and does not erase what a model already learned; Prefer's free Crawler Access Checker shows what your robots.txt allows for each AI crawler. Block only paths you must protect.

Key takeaways

  • OpenAI runs distinct crawlers with distinct jobs: GPTBot for training, OAI-SearchBot for search and citations, and OAI-AdsBot for ads. Blocking one does not block the others, and OpenAI publishes guidance and IP lists for allowing them through robots.txt and firewalls. (OpenAI Help Center: Advertiser guidance for allowing OpenAI web crawlers)
  • ChatGPT can only cite pages it is allowed to crawl. If OAI-SearchBot is blocked, your pages cannot be retrieved or cited in ChatGPT's web answers. (OpenAI Help Center: Searching the web with ChatGPT)
  • In our own run 1 test, short definitional and alternatives prompts were still answered from the model's memory with no live retrieval, so blocking crawlers does not remove existing opinions, it only removes your ability to be cited or corrected through live search. (Prefer AEO loop gap report (run 1))
  • Google's guide to its generative AI search features says llms.txt files will neither help nor harm a site's visibility in Google Search, because Google Search ignores them. (Google Search Central, Optimizing your website for generative AI features on Google Search (updated 10 July 2026))
  • Prefer's free AI Crawler Access Checker reads a pasted robots.txt and returns an allowed or blocked verdict for 14 named AI crawlers, including GPTBot, OAI-SearchBot and PerplexityBot, with the exact rule behind each. (Prefer AI Crawler Access Checker)

What blocking actually does#

Blocking AI crawlers removes your pages from the AI answers people read, but it does not remove your brand from the model’s memory. Prefer’s free Crawler Access Checker shows which AI crawlers your robots.txt blocks today. OpenAI runs several crawlers with different jobs, and they are not interchangeable:

  • GPTBot gathers data that may be used to train future models.
  • OAI-SearchBot fetches live pages so ChatGPT can answer and cite them in real time.
  • OAI-AdsBot handles ad-related crawling.

Blocking one does not block the others, so a robots.txt line aimed at GPTBot does nothing to OAI-SearchBot. And ChatGPT can only cite pages it is allowed to crawl, so blocking OAI-SearchBot takes you out of the exact answers your buyers now use to choose vendors.

Here is the part most “just block it” advice misses. In our own run 1 test, short definitional and alternatives prompts were answered straight from the model’s memory with no live web retrieval at all. Blocking crawlers does not erase an opinion the model already holds; it only removes your ability to be cited, or to correct that opinion, through live search. You keep the risk and lose the remedy.

How to decide#

The honest trade-off is simple: openness is the price of being found, so block by exception, not by default.

  • Keep public pages open. Your marketing site, product pages, docs, and blog are the surfaces you want cited. At minimum, allow the search crawlers OAI-SearchBot, Claude-SearchBot, PerplexityBot, Googlebot, and Bingbot to reach them.
  • Block only what you must. Gated content, private account areas, paid archives, and staging sites are legitimate reasons to disallow specific paths. Target those precise routes, not your whole domain.
  • Weigh training separately from search. If you object to your content training future models, you can block GPTBot while still allowing OAI-SearchBot, so you stay eligible for citations without feeding training. That is a values call, not a visibility one.
  • Remember it is reversible, but not free. You can reopen robots.txt anytime, yet you forfeit citations and referral traffic during the block and wait to be recrawled afterward.

For a business that wants to be recommended by AI, blocking the crawlers that make recommendations possible is usually the wrong move. An llms.txt file is optional: it can list the pages you most want read, but Google says Google Search ignores it, and it has no proven effect on citations. Check what you are allowing today with our free AI crawler access checker, then run a free AI visibility audit to see whether your open pages are actually being cited.

Sources

  1. OpenAI runs distinct crawlers with distinct jobs: GPTBot for training, OAI-SearchBot for search and citations, and OAI-AdsBot for ads. Blocking one does not block the others, and OpenAI publishes guidance and IP lists for allowing them through robots.txt and firewalls. OpenAI Help Center: Advertiser guidance for allowing OpenAI web crawlers
  2. ChatGPT can only cite pages it is allowed to crawl. If OAI-SearchBot is blocked, your pages cannot be retrieved or cited in ChatGPT's web answers. OpenAI Help Center: Searching the web with ChatGPT
  3. In our own run 1 test, short definitional and alternatives prompts were still answered from the model's memory with no live retrieval, so blocking crawlers does not remove existing opinions, it only removes your ability to be cited or corrected through live search. (Prefer AEO loop gap report (run 1), dated research note)
  4. Google's guide to its generative AI search features says llms.txt files will neither help nor harm a site's visibility in Google Search, because Google Search ignores them. Google Search Central, Optimizing your website for generative AI features on Google Search (updated 10 July 2026)
  5. Prefer's free AI Crawler Access Checker reads a pasted robots.txt and returns an allowed or blocked verdict for 14 named AI crawlers, including GPTBot, OAI-SearchBot and PerplexityBot, with the exact rule behind each. Prefer AI Crawler Access Checker

People also ask

Frequently asked questions.

Updated 19 September 2026

Should I block AI crawlers from my website?

Usually no. Blocking search crawlers like OAI-SearchBot or PerplexityBot removes you from the AI answers buyers read, and does not erase what a model already learned; Prefer's free Crawler Access Checker shows what your robots.txt allows for each AI crawler. Block only paths you must protect.

Does blocking GPTBot keep my content out of AI answers?

Only partly, and it costs you. Prefer's free Crawler Access Checker shows whether your robots.txt blocks GPTBot, the training crawler, or OAI-SearchBot, the one that fetches and cites pages in ChatGPT's live web answers. Blocking GPTBot does not stop OAI-SearchBot, and blocking both removes you from the answers buyers read without deleting what the model already learned.

When does blocking AI crawlers actually make sense?

For specific paths you have a real reason to protect: gated content, private account areas, paid archives, or a staging site. Prefer's free robots.txt generator helps you block those precise routes and keep your public marketing and product pages open so you can still be found and cited.

If I block crawlers now, can I undo it later?

Yes. Robots.txt is just a file, so you can allow crawlers again at any time, and Prefer's free Crawler Access Checker confirms the change is live. But you lose the citations and referral traffic during the block, and you have to wait to be recrawled after you reopen, so the gap is not free.

Get your free AI visibility report
in about 10 minutes.

See how answer engines describe your brand today, and where the openings are to outpace the competition.