Technical site audit, with history
Works todayFetches one page through the SSRF guard, runs seventeen on-page checks (the original ten — title, description, h1, image alt, canonical, noindex, thin content — plus seven more: lang attribute, heading-order breaks, a canonical pointing elsewhere, mixed content on an https page, vague link text, incomplete Open Graph tags, and duplicate hreflang), stores the result, and — from the second run of the same URL onward — reports exactly what changed, flagging regressions separately from fixes. A bounded, robots.txt-respecting whole-site crawl now runs alongside the single-page audit — capped at 100 pages and depth 4, run inline so a caller can retry without wondering whether the first one is still going — and closes the four gaps the single-page audit could never see on its own: broken internal links, redirect chains, orphan pages that are in the sitemap but linked from nowhere, and site-wide duplicate titles.
Will not: Crawl past its own cap. A crawl stops at 100 pages and 4 levels deep; past that, result.truncated comes back true rather than the crawl quietly running forever against someone else's server. Fetch anything Core Web Vitals or load-time related as part of an audit or a crawl — that is a separate capability (site-speed-and-core-web-vitals-monitor) needing a free PageSpeed Insights key.
Full capability page