RockethorseResearchBot

Crawler identification and contact page · rockethorseresearch.com

What this is

RockethorseResearchBot is an automated web crawler operated by the registrant of this domain. If you found this page from a User-Agent string in your server logs, you are in the right place: this page identifies the crawler, describes how it behaves, and tells you how to reach a person or exclude the crawler from your site.

Purpose

The crawler gathers publicly available information, primarily product and catalog pages, for private market and product research connected with prospective commercial projects. Retrieved pages are stored and analyzed internally. Page content is not republished, and access to the collected material is not sold or provided to third parties.

How it behaves

User-Agent string

RockethorseResearchBot (+https://rockethorseresearch.com/bot; contact@rockethorseresearch.com)

Excluding the crawler

To exclude it from your entire site, add this to your robots.txt:

User-agent: RockethorseResearchBot
Disallow: /

Path-specific Disallow rules and Crawl-delay are honored as well. Robots rules are rechecked regularly and a change is normally picked up within 24 hours; while a site’s rules are cached, the crawler will not knowingly act against a published change.

Contact and opt-out

contact@rockethorseresearch.com is a monitored inbox read by a person. Site operators may request exclusion by email instead of (or in addition to) robots.txt; such requests are honored and confirmed by reply, normally within ten business days.