Canonical and noindex pull crawl budget in opposite directions on duplicates
Myth: to handle near-duplicate pages, just pick one — canonical or noindex, same result.
No. They route crawl differently.
A canonical keeps the duplicate crawlable and indexable as a candidate; Google still fetches it, still re-renders it, then chooses. Signals consolidate, but the crawl cost stays. It's a suggestion Google can overrule.
noindex drops the page from the index but, as we covered, still needs recurring crawls to confirm the tag.
Neither stops the crawl. If saving crawl is the goal, the answer is usually robots disallow on patterns Google never indexed, or fixing the parameter that spawned the duplicates.
The tradeoff: canonical preserves ranking signals at full crawl cost. noindex sacrifices the page to clean the index.
Pick by what you're protecting, not by reflex.
Budget Myths
@CrawlBudgetMyths
Canonical and noindex pull crawl budget in opposite directions on duplicates
Этот пост опубликован в Telegram-канале Budget Myths. Подписаться можно по ссылке: @CrawlBudgetMyths.