What the crawler requests
The audit uses GET requests to inspect the submitted page, its public robots.txt, a sitemap and optional llms.txt. If the default sitemap is unavailable, it may try the first alternative sitemap declared in robots.txt. It follows up to five public HTTP or HTTPS redirects per resource, validates each destination, sets no Cookie header and does not execute page JavaScript or bypass access controls.
The user agent is CiteVisorAudit/1.0 (+https://citevisor.com/bot). Only parsed technical observations are retained in a private report. Response bodies are used during the inspection and are not saved.
Frequency and limits
Default limits are ten submissions per visitor per hour and 500 submissions across the service per UTC day. A normalized URL result is reused for 60 minutes to avoid another crawl when possible. Operators may lower these limits or pause audits. Each resource has a 12-second timeout, with a 50-second total audit budget and response size limits.
The audit is limited to one submitted page and its supporting files. It does not recursively crawl a website, send model requests or perform ongoing monitoring.
Website owner opt-out
If you operate a website and want its host excluded from this audit, contact contact@citevisor.com with the domain and enough information to establish ownership. The operator maintains a host blocklist and can add the host and its subdomains after reviewing the request.
We do not claim that this process is an instant self-service opt-out. Do not use this tool to submit a URL you have no right to inspect or a URL containing confidential information.
What a result means
Checks use the initial HTML and HTTP signals from a single request. They do not establish rendered browser content, actual AI crawler access, indexing, citations or rankings. Training and search permissions are described separately.
See the privacy notice for report retention and deletion and the terms of use for acceptable use.