Coverage and limits
Where this data is strong, where it is weak, and what it cannot see at all.
Our lists are smaller than the biggest vendors'. We would rather tell you that here than have you find it after buying.
What we hold
Where competitors beat us
Coverage. Vendors running their own crawler index more sites. On the technologies we have compared directly, we hold 46–65% of BuiltWith's top-1M counts. If completeness is what you need, buy theirs.
Freshness. One crawl a month, published 4–6 weeks after the month it covers. Nobody should use this for same-week triggers.
Large enterprises. Coverage falls sharply at the very top. Big organisations run marketing and commerce on subdomains a homepage crawl never sees. This data is strongest in the mid-market and thin above it.
Contacts. We do not sell contact records. We verify addresses you already hold.
Contact fields are sparse — here are the real rates
These come from each prospect's own published structured data, so they are free but patchy. Measured on the current crawl, not estimated:
A phone number on 3.9% of rows is not a phone column, and we do not present it as one. These rates ship inside every API response and CSV.
Feeds we withhold
A technology losing a large share of its install base every month is exhibiting detection failure, not customer behaviour. We compute implied churn against the whole crawl each month and withhold the feeds that fail it — 7 outright, 23 flagged. Mautic reads 265% monthly churn and Shopify Chat 174%, which is why neither is sellable as a feed however large its install base. The screen can only reject; passing it means nothing is provably broken, not that a feed is verified.
Contact coverage varies by segment
Named contacts are licensed, not observed in the crawl, and the source is strongest on US companies. One headline number would mislead in one direction or the other, so here is the whole distribution.
By country: 67.9% of US domains, 20.8% India, 16.4% United Kingdom, 9.9% Germany. A European list will carry far fewer named contacts than a US one, and the builder shows the count for the exact set you have selected before you pay for it.
What we cannot see at all
- Databases, warehouses, and server-side infrastructure — anything invisible to a rendered page.
- Sites with too little traffic to appear in Chrome's usage data.
- Technology behind a login, or on subdomains we do not crawl.
- Shared-hosting domains where thousands of sites collapse into one — we flag and exclude those, because a merged stack belongs to nobody.
Smaller, but from a source you can audit yourself and a history nobody else holds. That is the trade, stated up front. How the data is built →