Paste your domain. This checker reads your robots.txt and sitemap, then samples interior pages to find the places crawl budget leaks: URLs Google crawls but can't index, pages that quietly point their canonical elsewhere, parameter sprawl, and sitemaps full of dead or duplicate URLs. Add three numbers from Search Console and it will quantify the leak in real terms.
The live crawl finds the structural leaks. These three numbers turn them into a real percentage. Find them in Google Search Console → Indexing → Pages (indexed vs. not indexed) and → Settings → Crawl stats (total crawl requests). Leave any blank you don't have.
Runs entirely in your browser; no data is stored, and the Search Console numbers you type never leave the page. The crawl samples your sitemap and a handful of interior pages through a fetch proxy, so some sites will refuse the request. This tool detects the structural signals that predict wasted crawl budget — the definitive audit reads your server log files and Search Console crawl stats directly, which is the paid engagement this is a preview of. Built by a practitioner who does technical SEO for national brands.
What "crawl budget" means. Search engines only crawl so many URLs on your site in a given window. That budget is wasted when it's spent fetching pages that can't rank — duplicates, redirects, dead links, pages tagged "don't index" — instead of the pages you actually want found. On a small site it rarely matters; on a large or fast-growing one, waste means your important pages get crawled late or missed.
A blue i is context, not a grade — it tells you what the report is based on or what couldn't be checked from a browser.
The score reflects only what a live crawl can see (sitemap health + a sample of pages). It's a directional read, not a substitute for a log-file audit.
?sort=, ?filter=, tracking tags, etc. can multiply into thousands of near-identical URLs to crawl.The live crawl finds structural leaks; your Search Console numbers turn them into a real percentage. In Search Console: Indexing → Pages gives "indexed" vs "not indexed"; Settings → Crawl stats gives total crawl requests. A large "not indexed" share is your crawl leak, measured on your actual site.
Accuracy note: the report labels whether it ran at full accuracy (true redirect chains and header-level signals) or standard accuracy (via public proxies). Either way, the definitive picture of what search engines actually crawl lives in your server log files — that's the paid audit this previews.