Crawler transparency

HyperChatBot reads websites
with permission.

HyperChat uses public website content to build a private knowledge base requested by the website owner or an authorised service provider. It is not a general-purpose search crawler and it does not attempt to bypass access controls.

Crawler identity

Requests use this identifiable user agent:

HyperChatBot/1.0 (+https://hyperchat.uk/crawler)

When HyperChat crawls

A direct business customer starts a build by submitting its own website, or an agency records its client authority and verifies the domain by DNS, an HTTPS well-known file, or HyperChat's internal contract review. Agency builds cannot enter the queue while that authority is pending or revoked.

Technical limits

  • Only HTTP and HTTPS public-network URLs are accepted.
  • robots.txt is requested before page content and its applicable directives are enforced.
  • Crawling stays on the exact authorised hostname and refuses cross-host redirects.
  • Private, loopback, link-local and cloud-metadata network targets are rejected.
  • Page count, response time and extracted knowledge size are bounded.
  • HyperChat does not log in, solve access challenges or bypass paywalls.

Block or revoke access

Website owners can disallow the crawler in robots.txt:

User-agent: HyperChatBot
Disallow: /

An agency can also revoke a domain authority record, which pauses chatbots that rely on it. To report unauthorised use, email clone@clonecentre.ai with the domain and relevant evidence.

Content rights and personal data

Technical verification confirms control of a domain; it does not create copyright, confidentiality or data-protection rights. Customers and agencies remain responsible for holding all permissions and lawful bases required for their intended use. See our terms and privacy notice.