Skip to main content

Hominin

HominBot

HominBot is the web crawler for Hominin search. It fetches public web pages so they can appear in Hominin's search results.

Recognising it

Requests from HominBot carry a User-Agent like:

HominBot/1.0 (<crawler name>; +https://hominin.com/bot)

Controlling it with robots.txt

HominBot follows robots.txt (RFC 9309). It reads the group for User-agent: HominBot, or User-agent: * when there is none. To keep HominBot out of your whole site:

User-agent: HominBot
Disallow: /

What it collects

The text and title of a page, the links on it, and the addresses and descriptions of images on it. Pages behind a login are not crawled.

Contact

Questions about HominBot, or requests about how it crawls your site: homininglobal@protonmail.com.