OpenAI says why a page is missing from its index
The page describes a setting for Enterprise, Edu, Healthcare and regulated workspaces where ChatGPT answers from "OpenAI's indexed and cached web content" instead of live search. It lists, verbatim, the common reasons content is missing or stale: "The site blocks crawling through robots.txt or similar controls," "The site uses CDN or bot-blocking behavior," "The page requires login or personalization," "The content depends heavily on scripts or dynamic loading," and "The page is new, rarely accessed, or low-signal." It also says there is "no refresh SLA for a specific URL." We found it because Otterly's post today cites it.