AI glossary · G
AI glossary
GPTBot

What isGPTBot?

This crawler is where you decide whether your content should feed OpenAI's model training.

Get in touch
Enquiry
Nikolai Schöbel und Jeremias Burger, Co-Founder Scalableloops

Let's talk about your project.

First we check whether the project fits your business model. Then you get a proposal with phases and effort.

Check your AI visibility or call: +49 151 1576 5566
AI glossary · GPTBot

GPTBot is the crawler OpenAI uses to collect content for training its foundation models. It identifies itself by name, so you can block or allow it on its own in robots.txt. It has nothing to do with ChatGPT search, which OpenAI runs with a separate crawler.

In brief
  • According to OpenAI, blocking GPTBot indicates that your content should not be used to train its foundation models.
  • ChatGPT search relies on OAI-SearchBot, and user-triggered fetches run through ChatGPT-User.
  • OpenAI says it takes around 24 hours for its systems to reflect a change to robots.txt.
  • In our survey on 26 Sep 2026, none of 22 small and mid-sized business websites blocked GPTBot.
As of 10 Oct 2026Nikolai Schöbel and Jeremias Burger
Nikolai SchöbelJeremias Burger

Nikolai Schöbel and Jeremias Burger

Co-founders of Scalableloops GmbH. Nikolai Schöbel leads online marketing and AI strategy, Jeremias Burger the AI architecture.

On this page
  1. What OpenAI uses GPTBot for and why your ChatGPT visibility does not depend on it
  2. What belongs in robots.txt for OpenAI's GPTBot without blocking more than you intended
  3. Frequently asked questions
  4. Where to go deeper
  5. Terms you should know in the same context
Purpose

What OpenAI uses GPTBot for and why your ChatGPT visibility does not depend on it

OpenAI splits three jobs across three crawlers. GPTBot gathers training material, OAI-SearchBot builds the index for ChatGPT search, and ChatGPT-User fetches a page when someone asks for it in a chat. Each one uses its own name, the user agent, and gets its own rules in robots.txt.

That means you can say no to training without disappearing from the answers. As long as OAI-SearchBot is allowed, your pages can still be found in ChatGPT search, even with GPTBot locked out. How the other providers split their crawlers is covered under AI crawler.

Setup

What belongs in robots.txt for OpenAI's GPTBot without blocking more than you intended

The rule takes two lines: User-agent: GPTBot followed by Disallow: /. Be careful if you also block directories for all crawlers and create separate groups for other bots. Under the RFC 9309 standard, a crawler follows only the group that matches its name most precisely and ignores all others, including User-agent: *.

A block like this is a request, not protection. robots.txt cannot enforce anything, and anything truly confidential belongs behind a password. Check your firewall and bot protection as well, because they sometimes turn crawlers away without robots.txt showing any trace of it.

Frequently asked questions

Frequently asked questions

What is GPTBot?

The crawler OpenAI uses to collect content for training its foundation models. Disallowing it signals that your content should not be used for that training.

What happens if I block GPTBot?

According to OpenAI, your content will then not be used to train its foundation models. Your visibility in ChatGPT search does not depend on it, because OAI-SearchBot handles that.

Do many companies block GPTBot?

Hardly any so far. In our non-representative survey of 22 small and mid-sized business websites on 26 Sep 2026, none blocked GPTBot, mostly because their robots.txt did not mention it at all.

Related terms

Terms you should know in the same context

OAI-SearchBot · AI crawler · Google-Extended · llms.txt · AI visibility · Back to the AI glossary A to Z

Blocking GPTBot is a decision about training, not about your visibility.

or call: +49 151 1576 5566

Projekt-Detail

    Got a project in mind?

    We reply personally. First a use-case check, then an architecture proposal.

    Start your inquiry