What blocking actually does#
Blocking AI crawlers removes your pages from the AI answers people read, but it does not remove your brand from the model’s memory. Prefer’s free Crawler Access Checker shows which AI crawlers your robots.txt blocks today. OpenAI runs several crawlers with different jobs, and they are not interchangeable:
- GPTBot gathers data that may be used to train future models.
- OAI-SearchBot fetches live pages so ChatGPT can answer and cite them in real time.
- OAI-AdsBot handles ad-related crawling.
Blocking one does not block the others, so a robots.txt line aimed at GPTBot does nothing to OAI-SearchBot. And ChatGPT can only cite pages it is allowed to crawl, so blocking OAI-SearchBot takes you out of the exact answers your buyers now use to choose vendors.
Here is the part most “just block it” advice misses. In our own run 1 test, short definitional and alternatives prompts were answered straight from the model’s memory with no live web retrieval at all. Blocking crawlers does not erase an opinion the model already holds; it only removes your ability to be cited, or to correct that opinion, through live search. You keep the risk and lose the remedy.
How to decide#
The honest trade-off is simple: openness is the price of being found, so block by exception, not by default.
- Keep public pages open. Your marketing site, product pages, docs, and blog are the surfaces you want cited. At minimum, allow the search crawlers OAI-SearchBot, Claude-SearchBot, PerplexityBot, Googlebot, and Bingbot to reach them.
- Block only what you must. Gated content, private account areas, paid archives, and staging sites are legitimate reasons to disallow specific paths. Target those precise routes, not your whole domain.
- Weigh training separately from search. If you object to your content training future models, you can block GPTBot while still allowing OAI-SearchBot, so you stay eligible for citations without feeding training. That is a values call, not a visibility one.
- Remember it is reversible, but not free. You can reopen robots.txt anytime, yet you forfeit citations and referral traffic during the block and wait to be recrawled afterward.
For a business that wants to be recommended by AI, blocking the crawlers that make recommendations possible is usually the wrong move. An llms.txt file is optional: it can list the pages you most want read, but Google says Google Search ignores it, and it has no proven effect on citations. Check what you are allowing today with our free AI crawler access checker, then run a free AI visibility audit to see whether your open pages are actually being cited.
Sources
- OpenAI runs distinct crawlers with distinct jobs: GPTBot for training, OAI-SearchBot for search and citations, and OAI-AdsBot for ads. Blocking one does not block the others, and OpenAI publishes guidance and IP lists for allowing them through robots.txt and firewalls. OpenAI Help Center: Advertiser guidance for allowing OpenAI web crawlers
- ChatGPT can only cite pages it is allowed to crawl. If OAI-SearchBot is blocked, your pages cannot be retrieved or cited in ChatGPT's web answers. OpenAI Help Center: Searching the web with ChatGPT
- In our own run 1 test, short definitional and alternatives prompts were still answered from the model's memory with no live retrieval, so blocking crawlers does not remove existing opinions, it only removes your ability to be cited or corrected through live search. (Prefer AEO loop gap report (run 1), dated research note)
- Google's guide to its generative AI search features says llms.txt files will neither help nor harm a site's visibility in Google Search, because Google Search ignores them. Google Search Central, Optimizing your website for generative AI features on Google Search (updated 10 July 2026)
- Prefer's free AI Crawler Access Checker reads a pasted robots.txt and returns an allowed or blocked verdict for 14 named AI crawlers, including GPTBot, OAI-SearchBot and PerplexityBot, with the exact rule behind each. Prefer AI Crawler Access Checker
People also ask
- Should I block GPTBot in my robots.txt?
- Is it a good idea to stop AI from crawling my website?
- Will blocking AI crawlers hurt my visibility in ChatGPT?