← AI crawlers
OpenAI · Model training

GPTBot

GPTBot collects content for training OpenAI's models. It isn't what ChatGPT uses to find businesses.

THE FACTS
Operated by
OpenAI
robots.txt token
GPTBot
User-Agent sent
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko); compatible; GPTBot/1.4; +https://openai.com/gptbot
Honours robots.txt
Yes — OpenAI documents that it honours robots.txt.
IF YOU BLOCK IT

This one is not on the answer path.

GPTBot collects content for training OpenAI's models — it isn't what ChatGPT uses to find businesses, so this doesn't affect whether you show up in ChatGPT's answers.

Whether to allow model training is a judgement call about your content, not a visibility decision — so this page does not make a recommendation either way.

THE EXACT LINES

Two lines in robots.txt.

Block GPTBot

User-agent: GPTBot
Disallow: /

Goes in the robots.txt at the root of your domain.

Allow GPTBot

User-agent: GPTBot
Allow: /

Only needed if a broader rule already blocks it — a group of its own overrides the wildcard.

Group selection matches the token, not the full User-Agent string, and it is case-insensitive. A rule under User-agent: * applies only when the crawler has no group of its own.

DON’T CONFUSE IT WITH

OpenAI runs more than one, and they do different jobs.

Blocking one does not block the others. This is where most advice goes wrong: the rule that stops model training is not the rule that decides whether you appear in an answer.

SOURCE

Everything on this page was read from OpenAI’s own published source on . Read the source. Vendors change these pages without announcing it — if you find something here that no longer matches, tell us and we’ll correct it.

CHECK YOUR OWN SITE

Knowing what GPTBot does doesn’t tell you whether your site lets it in.

The free checker reads your robots.txt and asks your site for its homepage as each major AI crawler, then tells you which ones it turned away. No signup, about 20 seconds.