Inventory the surfaces
Separate public pages, resources needed for rendering, private areas and variants without independent value.

GUIDE / PROTOCOL
A clear access policy starts with the intended use of each public page rather than a list of crawler names.
Direct answer
Decide which pages may be discovered, which may appear in search and which should be excluded from a specific use. Express those choices through robots.txt, page directives and server responses, then test what a crawler can actually read.
01 / Review method
Separate public pages, resources needed for rendering, private areas and variants without independent value.
Decide access for conventional search, AI search and training collection using the crawler identities documented by each operator.
Test HTTP status, robots.txt, redirects, robots meta or X-Robots-Tag and served HTML, including language versions.
Compare server logs, webmaster tools and referral traffic. A crawler visit proves neither indexing nor citation.
02 / Evidence to keep
Keep the version, date and reason for each robots directive.
Record the exact page, status, canonical and page-level rules.
Separate declared search, preview and training crawlers.
Document what remains accessible and what actually appears in the available tools.
03 / FAQ
No. It enables access needed for possible appearance in ChatGPT search but does not guarantee use of a page.
Yes when an operator documents separate agents. Verify names and behaviour in official documentation before changing the rules.
04 / Official references