Most affiliate teams carry the ideal partner in their head, not on paper. We convert that instinct into a testable signal set, calibrate it against a reviewed slice, and measure the true count across 120M+ classified domains — so you plan against a real number instead of a guess.
A prose description of your ideal partner cannot be run against the web. It has to become a set of signals a machine can extract and a human can audit. Two questions block every recruitment plan until that happens.
Which signals are non-negotiable, which merely help, and which quietly disqualify a site? Until that is explicit, every reviewer applies a slightly different bar and the list drifts.
Is the addressable set 4,000 publishers or 140,000? The answer decides headcount, budget, tooling, and whether the channel is worth building at all — yet most teams have never measured it.
Each step is explicit and inspectable. You see the draft definition, the reviewed slice, the measured error, and the extrapolated size — before a single dollar of budget is committed.
In a working session we translate your "ideal publisher" into required, preferred, and disqualifying signals drawn from a fixed vocabulary — plus any custom fields your program needs. The definition is written down first.
We draw a representative ~1,000-domain random slice from your vertical inside the 120M+ classified-domain database and screen it with the draft ICP. First-pass output is ready in hours, not weeks.
Your team marks each sampled domain fit / not-fit. Disagreements are the signal: they show which definitions to sharpen. We adjust wording and thresholds and re-screen the slice.
The loop repeats until measured precision clears your bar — in the reference engagement the reviewed sample landed at ~96% precision, roughly a 4% error rate. The number is measured, not asserted.
The fit rate on the reviewed slice, applied to the full category count, yields the addressable publisher estimate — labeled as an estimate, with the calibration basis shown alongside it.
You leave with a documented signal set, a measured precision figure, and a full-run yield projection — the inputs a recruitment plan actually needs.
An illustrative loop for a mid-market advertiser recruiting independent content publishers. The slice figures are the measured basis; the full-run figure is an extrapolation, labeled as an estimate.
The first pass over the calibration slice flagged the usual keyword-not-intent errors — competitor storefronts that read as "relevant", and templated affiliate pages with no real editorial. Reviewed precision started well below the bar.
We tightened the "independently run" and "editorial structure" definitions and hardened the own-inventory disqualifier. Each re-screen moved measured precision up and cut the false-positive rate.
Once the reviewed slice held at ~96% precision (~4% error), the fit rate was applied to the full category count to project the addressable publisher universe — reported as an estimate with the slice as its basis.
In a large-scale production run on our own classification platform, the same define-calibrate-size loop ran across the travel vertical. The client reviewed a sample of the output against their own judgment.
Reference: a large-scale production run on our own classification platform. Your own sizing figures come from your ICP and your reviewed slice, not from this run.
Any publisher profile we show on a public page is drawn from a real screening run, so we follow standard confidentiality practice: domains are withheld and descriptions generalized. The published profiles are therefore intentionally not traceable to the sites — including through a web search — which protects the publishers without changing the underlying data. Your pilot and client deliverables contain the actual domains with every signal, so everything can be verified directly. Start with a free pilot →
You receive the documented ICP, a per-domain deduplicated CSV, and a sizing sheet. Below is an illustrative slice of the CSV — one row per canonical domain, with the fit verdict and per-field confidence. Values are illustrative; real domains ship in your deliverable.
| canonical_domain | is_icp | fit_conf | independently_run | affiliate_links | fresh_content | own_inventory | language |
|---|---|---|---|---|---|---|---|
| publisher-a.example | true | 0.94 | 0.91 | yes | yes | no | en |
| publisher-b.example | true | 0.88 | 0.79 | yes | partial | no | en |
| publisher-c.example | false | 0.72 | 0.40 | no | yes | yes | de |
| publisher-d.example | true | 0.81 | 0.85 | partial | yes | no | fr |
The free pilot runs the full pipeline on your vertical, calibrates your ICP against a reviewed slice, and delivers your first 20 qualified publishers plus a full-run size projection. No cost, no obligation.
Request a Free Pilot