getdomaindata
getdomaindata is the crawler for getdomaindata.com, a service that compiles technology, hosting and infrastructure profiles of websites and sells access to that dataset. Its bot page states that it visits only the homepage of publicly reachable sites over ports 80 and 443, that page content is not retained, that it honors robots.txt for the user agent getdomaindata or for *, and it offers an exclusion form together with optout@getdomaindata.com and abuse@getdomaindata.com. No operating company, legal entity or country is disclosed anywhere on the site, the product is marked as coming soon, no IP ranges or reverse DNS hostnames are published, and the token does not appear in the public bot directories that were checked, so the stated purpose rests on the operator's own description alone. Traffic observed on this platform does not match that homepage-only claim: across a seven day window 43 percent of its requests went to paths other than a site root, including article and category pages.
At a glance
- Operator: getdomaindata.com
- Type: Scraper
How Centinel checks it
- User agent: The request calls itself this crawler. Anyone can send the same string.
getdomaindata publishes nothing Centinel can check a source against, so a match reports the name and leaves the source unconfirmed. The match tokens, verification domains, and address feeds are not published here.
Allowing or blocking it
The crawler object in the /validate response sets access_allowed to true only for a verified source that your tenant allowlists. A policy rule can allow or block this crawler by its category, Scraper.