Our crawler
CiteRep makes automated HTTP requests only after an authenticated user submits a site for audit or enables scheduled monitoring. Requests identify the CiteRep audit service and are limited by the report's page, time, and cost scope.
What we request
We may request the submitted public URL, same-host public links, robots.txt, sitemap files, llms.txt, and related public assets needed for analysis. We do not intentionally sign in, solve access challenges, or bypass paywalls or technical restrictions.
Safety controls
- Only HTTP and HTTPS destinations are accepted.
- Private, loopback, link-local, and reserved network targets are blocked.
- Report types impose fixed page and AI-call limits.
- Every user-triggered report must consume a daily free entitlement or an available paid report token.
- Monthly monitoring is limited to entitled sites and a fixed cadence.
- Failed or timed-out jobs stop and restore the applicable entitlement.
Stored audit data
We store page URLs and extracted metadata, findings, scores, recommendations, model usage, and report history in the workspace. Selected public webpage text may be included in system-generated AI prompts. AI call traces are compressed in object storage and scheduled for deletion within 90 days.
Deletion and objections
Workspace members can request deletion of a site, report, or account through support. A site owner who believes CiteRep accessed a site without authority can email support@tensabyte.com. Include the domain, relevant logs if available, and a method for us to verify your relationship to the site.
No training sale
We do not sell crawled content or use customer audit content to train a general-purpose advertising profile. External AI providers process selected content to return the requested analysis under their service terms and our configured data controls.