Our crawler

The getdomaindata.com crawler

getdomaindata.com operates an automated crawler (which identifies itself asgetdomaindata/1.0) that visits the homepage of publicly reachable websites to detect the technologies and infrastructure they run. This page explains what it does and how to opt out.

How to identify it

Every request our crawler makes carries this User-Agent:

Mozilla/5.0 (compatible; getdomaindata/1.0; +https://getdomaindata.com/bot)

What it does

What it does not do

Opt out

The simplest way to opt out is your robots.txt: disallow the User-agent getdomaindata (or *) and our crawler will not request your pages.

You can also have a domain, hostname, or IP range permanently excluded — email [email protected] with the domains or network ranges you want removed. We add exclusions promptly and they persist across future crawls.

Network operators & abuse reports

If you operate a network and observed traffic from our crawler, please reach out to [email protected]. We honor exclusion requests, maintain a permanent do-not-contact list for reported sensor and honeypot ranges, and are happy to coordinate on attribution or whitelisting.

Operated by getdomaindata.com. Abuse & opt-out contact: [email protected].