OpenAI
OAI-SearchBot
Discovers content for ChatGPT search results.
AI searchCollects web content that may be used to train OpenAI foundation models.
See your agent trafficTHE SIGNAL
A crawl indicates content collection, not a ChatGPT recommendation or a prospective customer. Keep this series separate from user-requested visits.
Choose your training policy independently of search visibility. Review public URLs and response codes before changing access.
Look for the GPTBot identifier in a request’s User-Agent. This is a name match, not identity verification. Version strings may change.
Compare the request IP with the current operator-published ranges linked in the source. A matching name alone does not verify identity.
Use the GPTBot robots.txt group to control training collection. This setting is independent of OAI-SearchBot search access.
Optional full-site opt-out. Merge with existing rules only if intended. robots.txt does not secure private content.
User-agent: GPTBot
Disallow: /403 means access was denied; 404 means the resource was not found. Compare the public path, request time, edge security event and origin response to find the cause.
Separate intended restrictions and secret-file probes from pages that should work. A User-Agent name alone does not justify allowing a request.
Identity and purpose are based on these sources. Analytics interpretation and suggested checks are Apostl guidance.
See their requests. Find the pages that matter.
See your agent traffic