Our crawler
getdomaindata.com operates an automated crawler (which identifies itself asgetdomaindata/1.0) that visits the homepage of publicly reachable websites to detect the technologies and infrastructure they run. This page explains what it does and how to opt out.
Every request our crawler makes carries this User-Agent:
Mozilla/5.0 (compatible; getdomaindata/1.0; +https://getdomaindata.com/bot)robots.txt — if your site disallows our crawler (User-agent getdomaindata, or *), we will not request its pages.The simplest way to opt out is your robots.txt: disallow the User-agent getdomaindata (or *) and our crawler will not request your pages.
You can also have a domain, hostname, or IP range permanently excluded — email [email protected] with the domains or network ranges you want removed. We add exclusions promptly and they persist across future crawls.
If you operate a network and observed traffic from our crawler, please reach out to [email protected]. We honor exclusion requests, maintain a permanent do-not-contact list for reported sensor and honeypot ranges, and are happy to coordinate on attribution or whitelisting.