Acquisition Universe doesn't search a stale list. It runs against effectively the entire active web: 120 million domains classified into the full IAB taxonomy of 700+ categories, re-checked monthly, with newly-registered domains added daily — built on the same enterprise categorization infrastructure behind Alpha Quantum's commercial databases.
Alpha Quantum is a leading provider of website categorization APIs and offline (downloadable) categorization databases — enterprise-grade classified datasets trusted by 300+ clients for content filtering, brand safety, cybersecurity and market intelligence. Acquisition Universe is that same production infrastructure and those same leading enterprise databases, pointed at acquisition screening. You're not getting a bolt-on scraper; you're plugging into a categorization platform that already operates at web scale.
This is standing infrastructure, not a one-off script. The web changes constantly, so the universe is rebuilt and re-verified on a schedule.
The full domain universe categorized against the 700+ category IAB taxonomy — the schema itself rides in every classification, so each domain is judged against the whole taxonomy, not a handful of buckets.
All 120M domains are re-classified every month. Sites go dark, pivot, rebrand and change owners; a database that isn't re-verified quietly rots. Ours is current to the month it ships.
Newly-registered domains (~300k/day across the web) are crawled, filtered and classified as they come online — so a target that launched last week is already in the universe.
On top of the base pass, 10 million domains a month go through a multi-step re-classification — several model passes with cross-checks — to lift precision where it matters most.
Suppose you decided to reproduce this instead of licensing it. A single monthly cycle would involve all of the following, at web scale — and the total swings hugely with the quality of LLM you choose: a budget model is cheap per call but weak on precision, while frontier models classify far more accurately at several times the cost.
Reading each domain and classifying it against the full 700+ category IAB taxonomy, with a confidence score on every result.
Several model passes with cross-checks on the highest-value domains, beyond the single base classification.
Fetching 120M+ live pages every month plus the daily new-domain intake — residential/datacenter proxy pools, retries, block-handling and the bandwidth behind them.
The cluster that fetches, extracts text and drives the classification pipeline end to end.
A 100M+ row classified database with history, kept queryable and safe.
Every sample on this site is one thin slice of that 120M-domain universe, screened against a specific acquisition thesis — with the evidence behind every signal.
Browse the samples Run your thesis — free pilot