ChatGPT-User
ChatGPT-User is a bot operated by OpenAI that fetches single pages on demand for an AI assistant. ChatGPT-User is OpenAI's web crawler that visits websites when ChatGPT users request information. It honours robots.txt, and a request claiming to be ChatGPT-User can be checked against a published IP range file.
ChatGPT-User at a glance
| User agent | Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko); compatible; ChatGPT-User/1.0; +https://openai.com/bot |
|---|---|
| Operator | OpenAI |
| Purpose | Assistant fetchers |
| Respects robots.txt | Yes, documented |
| Verification method | Published IP range file Fetches are triggered by a person asking ChatGPT about a URL, so traffic is spiky rather than a steady crawl. The egress ranges are published as JSON. https://openai.com/chatgpt-user.json |
| Crawl pattern | Only when prompted by a user. |
| Robots.txt tokens | ChatGPT-User |
| Operator documentation | https://platform.openai.com/docs/bots |
ChatGPT-User is OpenAI's web crawler that visits websites when ChatGPT users request information. This enables ChatGPT to include links in its responses.
Allow or block ChatGPT-User
Paste one of these into the robots.txt file at the root of your domain. Rules are per token, so a block on one crawler leaves every other bot untouched.
User-agent: ChatGPT-User
Disallow: /Blocks every path for this crawler only.
User-agent: ChatGPT-User
Allow: /Explicit allow. Useful when a wildcard rule above it would otherwise catch this crawler.
User-agent: ChatGPT-User
Allow: /
Disallow: /account/
Disallow: /checkout/
Disallow: /searchEdit the Disallow paths to match your own account, checkout and search URLs.
What blocking actually costs you
Blocking ChatGPT-User means a person who asks about one of your URLs gets an error instead of a summary. That is a support and reputation trade-off, not a training one.
Block every crawler in the directory at onceIs that really ChatGPT-User?
A user-agent header is a string the client chooses. Scrapers copy ChatGPT-User precisely because site owners allow it. Fetches are triggered by a person asking ChatGPT about a URL, so traffic is spiky rather than a steady crawl. The egress ranges are published as JSON.
Paste the IP address from your access log below. You will see whether it belongs to a hosting provider, a residential proxy pool or a Tor exit, plus the ASN that announces it, which is what tells you whether the claim holds up.
Other crawlers run by OpenAI
Related crawlers
Identify crawlers automatically instead of by hand
Agentscan checks a request against published crawler ranges, reverse DNS and hosting data, then returns a verdict your edge can act on. One call per request, no range files to keep up to date.