ChatGPT visibility,
bounded by what is documented
OpenAI documents four separate agents, not one, and they do different jobs: a search index crawler, a training crawler, an ads validator and a user-triggered fetcher. Confusing them is the most common mistake made about ChatGPT visibility, because blocking the training crawler and blocking the search crawler have completely different consequences. What OpenAI does not publish is how ChatGPT decides which sources to name — so OG01 measures what it actually returns about your business rather than theorising about the model.
Four agents, four different decisions
These are OpenAI's own descriptions, summarised. The distinction that matters commercially is between the agent that feeds ChatGPT's search answers and the agent that gathers training data, because a great many sites blocked both when they meant to block one.
- OAI-SearchBot — indexes sites for ChatGPT's search results. OpenAI states that a site opted out of this agent will not appear in ChatGPT search answers. If you want to be findable in ChatGPT, this is the one that matters.
- GPTBot — gathers content to train generative models. Disallowing it signals that your content should not be used for model training. That is a separate commercial decision from search inclusion, and it is the one most businesses actually mean when they say they blocked OpenAI.
- ChatGPT-User — fetches a page when a person asks ChatGPT to visit it. OpenAI states robots.txt rules may not apply here, because the request originates with a user rather than with automated crawling.
- OAI-AdsBot — validates landing pages submitted as advertisements, and visits only pages explicitly submitted rather than crawling the open web.
Why there is no factor list on this page
OpenAI publishes how its agents reach the web. It does not publish how ChatGPT selects which sources to cite, in what order, or with what weighting — and nobody outside the company can verify a list claiming otherwise. Any page offering you "the ChatGPT ranking factors" is offering a theory dressed as documentation.
What can be established from outside is narrower and considerably more useful: whether the conditions any retrieval-based system depends on are present on your domain. Can your pages be fetched and parsed? Does your business resolve to one clear entity? Are your claims corroborated anywhere other than your own site? And, most directly, what does ChatGPT actually say when asked about your business? That last one is an observation, repeatable and recordable, which is why OG01 runs prompts and reports the responses rather than estimating a score for a black box.
Conditions you control, and responses you can read
No part of this promises a citation. Responses vary between runs, and OG01 presents them as observations rather than as stable positions — because they are not stable, and a vendor charting them as though they were is selling you a graph of noise.
Where the platform facts on this page come from
Every statement above about OpenAI's agents is drawn from OpenAI's own documentation, summarised rather than reproduced. Reviewed 2026-09-08.
- OpenAI — Bots / crawler documentation — https://developers.openai.com/api/docs/bots
Vendors change documentation without notice. If something here no longer matches the source, the source wins and this page is wrong — tell us and it gets corrected.
ChatGPT visibility questions
If I block GPTBot, do I disappear from ChatGPT?
Can OG01 get my business cited by ChatGPT?
Does robots.txt control everything OpenAI does?
Why does the same prompt give different answers on different days?
Is there structured data that makes ChatGPT prefer my site?
A free reading returns in roughly a minute and a half, without a card. The deeper reading carries the prompts and the responses themselves.