Passive Site Footprint
List every URL of a domain that web archives have seen, without ever contacting the site: Common Crawl and the Wayback Machine, merged and sorted into documents, backups, config files, scripts, admin paths and parameters.
Target domain
Passive only: the target site is never contacted. Only the Common Crawl index and the Wayback Machine are queried, and every link opens the archived copy, not the live site.