Identify the crawler
Use published documentation to identify the user agent and its stated purpose before changing a control, and record the documented token alongside any change.
AI crawlers · technical reference
Crawler controls depend on the crawler and purpose. Separate search discovery from training-related access rather than treating every AI crawler as interchangeable.
Direct answer
Read the published crawler documentation and apply controls to the specific user agent and purpose. A control is not a guarantee of how any answer experience will use or cite content.
Reference guide
Use published documentation to identify the user agent and its stated purpose before changing a control, and record the documented token alongside any change.
OpenAI guidance distinguishes OAI-SearchBot for search discovery from GPTBot training controls; do not collapse those purposes.
Robots directives are crawler access guidance. They do not promise crawling, inclusion, sources, or citations.
Allowing or disallowing a documented user agent changes what compliant crawlers may request. It does not determine whether content is used in an answer, cited as a source, or summarized, and it cannot verify crawler identity by itself: pair directives with published documentation and ordinary server logs.
Primary sources
OpenAI. Platform guidance can change; recheck the source before acting.
OpenAI. Platform guidance can change; recheck the source before acting.
Related guides
Return to the educational GEO reference hub.
Write precise, documented directives for named crawlers.
Interpret returned references and crawler purposes in context.
Questions
No. Controls and meanings depend on the documented crawler and purpose.
Have a real research question?
Use supported workflows and keep their boundaries visible.
Start free trial