KenapticBot

About our crawler.

Kenaptic customers connect content sources they choose; KenapticBot fetches those pages to discover cross-link relationships. It stores only minimal derived metadata (URL, title, topic digest, embedding), republishes nothing, and its output is a plain hyperlink to your page.

Personal data is removed before storage

KenapticBot strips personal identifiers from every page it reads — email addresses, phone numbers, @handles, usernames, bylines and profile links — from the title, body and description, at the moment of ingestion and before anything is written down. Everything derived afterwards is built from the cleaned text, so no personal data is stored and none is ever sent to a language model.

This applies to every source we read, including sites operated by the customer who asked us to read them. Detection is plain pattern matching that runs on our own infrastructure: we do not send your text to a model in order to find the personal data in it.

This is data minimisation applied to everything, not a guarantee that pattern matching catches every possible form of personal data in free prose. If you would rather a page were not read at all, the opt-out below is immediate.

Identification

User-Agent: KenapticBot/1.0 (+https://www.kenaptic.com/bot)

How to opt out (self-service, immediate)

robots.txt is our opt-out mechanism. Block KenapticBot and we will not crawl your site; any existing Kenaptic-injected links pointing at your pages are automatically retracted on the next daily pass.

# robots.txt — block Kenaptic entirely User-agent: KenapticBot Disallow: /

We also honour Crawl-delay, machine-readable TDM reservations (tdmrep.json), and noai/notrain robots directives. KenapticBot never crawls behind logins and never bypasses technical access controls.

Contact

Questions or removal requests: bot@kenaptic.com.