GEOChecker
Data Sources
The checker uses public pages and public metadata from the submitted site.
Included
The crawler reads homepage HTML, robots.txt, sitemap URLs, sampled public pages, metadata, links, headings, text, and JSON-LD.
Excluded
It does not crawl login-only pages, paywalled content, personal data pages, or paths blocked by robots.txt.