How the monetization scanner works
1. Crawl & observe
The crawler respects robots.txt, samples prioritized public pages, classifies page intent and extracts outbound links, merchant destinations, social profiles and known technology fingerprints.
2. Evidence confidence
Direct link parameters and exact tracking hosts can receive high confidence. Inferred business methods receive lower confidence. Unknown signals are stored as learning candidates rather than named as facts.
3. Archetype & rule path
Sites receive weighted traits such as affiliate publisher, DTC ecommerce, multichannel merchant, community publisher, ad-supported publisher, SaaS, lead generation or marketplace. The strongest evidence determines the primary rule path.
4. Human-approved learning
New affiliate hosts, plugins and methodologies enter a management review queue. They only become production knowledge after approval. Approved knowledge increments the fingerprint/ruleset version and can queue existing sites for re-audit.
5. SEO quality gate
Public audit pages only become indexable when enough site-specific evidence has been collected. Thin, failed or low-confidence scans remain noindex.
6. Owner data is separate
Public observations never imply access to private revenue. Domain owners can verify control and later connect first-party analytics/affiliate sources for private, financially grounded insights.