Skip to content
BrandWater AI
Research note

Does blocking GPTBot affect what ChatGPT cites? A documentation-based analysis

4 October 2026 2 min read

Headline findings

  • OpenAI names three crawlers with three distinct, documented purposes.
  • Blocking GPTBot addresses training use. It does not, per the documentation, address search indexing or live fetching.
  • A page can still be found and cited in search-grounded answers while opted out of training.

Quick answer

OpenAI documents GPTBot for training, OAI-SearchBot for search indexing, and ChatGPT-User for live, user-triggered fetches. Blocking only GPTBot in robots.txt stops training use but, per OpenAI's own documentation, does not stop the other two, so it should not by itself remove a page from ChatGPT's search-grounded answers.

What OpenAI documents about each crawler

Documented purpose of each crawler (as published, accessed Sep 2026)
CrawlerDocumented purposeWhat blocking it in robots.txt is documented to do
GPTBotCrawls content that may be used to train OpenAI's modelsOpts the page out of training use
OAI-SearchBotCrawls and indexes pages for ChatGPT's search featuresKeeps the page out of that search index
ChatGPT-UserFetches a specific page live, triggered by a user's request or action inside ChatGPTPrevents that live, in-the-moment fetch

Why this gets confused

A common assumption is that one 'block ChatGPT' rule in robots.txt covers all of OpenAI's access. OpenAI's own documentation separates the three purposes and names three different user agents, so a rule written for GPTBot alone does not, per that documentation, reach the other two. A site that wants to opt out of training while staying visible in ChatGPT's search answers should treat these as three separate decisions, not one.

Method note

This reflects only what OpenAI's bots documentation states on the access date above. It does not measure actual citation behaviour, and OpenAI can rename or change crawler behaviour at any time.
Does blocking all three guarantee a page never appears in ChatGPT?

It should prevent OpenAI's documented access methods from reaching the page, but no vendor documentation offers a guarantee covering every possible path an answer could reference a page.

Is this the same pattern for other AI vendors?

Perplexity and Google document a similar separation between training, indexing and live fetching. See the companion piece on crawler types for the full comparison.

Sources

  1. 1. OpenAI: Overview of OpenAI crawlers (Accessed Sep 2026)

Cite this page

BrandWater AI Research. (2026, 4 October 2026). Does blocking GPTBot affect what ChatGPT cites? A documentation-based analysis. https://brandwaterai.in/research/does-blocking-gptbot-affect-chatgpt-citations

How we work

Figures are dated and linked to their sources. Where none exist we say so. Read our methodology and AI transparency pages.

Put the theory to work on your brand.

Start with a first audit. We show where you appear, who is named instead, and what to fix first.