Hunting orphan pages that quietly never get crawled
Orphans — pages with no internal links — are the real "not crawled" cases people blame on budget. Find them:
— Export all URLs from your sitemap/CMS database.
— Crawl your site with a desktop crawler (link-following only, no sitemap seed).
— Diff the two lists. URLs in the database but NOT in the crawl = orphans.
— Cross-check against logs: orphans usually show zero or near-zero Googlebot hits.
— Fix by adding real internal links from relevant, crawled pages — or delete them if they're junk.
Actually "Crawled — currently not indexed" and "Discovered — not indexed" often trace back to weak internal linking, not a starved budget.
If nothing links to a page, Google has no reason to crawl it. That's not a budget problem. It's a map problem.
Budget Myths
@CrawlBudgetMyths
Hunting orphan pages that quietly never get crawled
Этот пост опубликован в Telegram-канале Budget Myths. Подписаться можно по ссылке: @CrawlBudgetMyths.