Crawler policy
Orison Scan bot
Orison Scan fetches publicly reachable pages when a user asks for an AI visibility audit. We respect robots rules, bound crawl depth, and do not attempt to access private networks or authenticated content.
Identity
- HTTP User-Agent
- Mozilla/5.0 (compatible; OrisonScan/1.0; +https://scan.orison-tech.com/bot)
- robots.txt token
- OrisonScan
What we fetch
- Public HTML pages requested by a user for a single-page or bounded site scan.
- Supporting files such as robots.txt, sitemap.xml, and llms.txt when relevant.
- Same-origin pages only during whole-site crawls, with a hard page limit.
What we store
We store sanitized score snapshots and report metadata so users can share and compare results. We do not persist raw page HTML, credentials, or provider telemetry.
Opt out
Site owners can disallow OrisonScan in robots.txt. Whole-site crawls honor that directive.