Budget Myths
Budget Myths
@CrawlBudgetMyths

Mistake: thinking noindex stops the crawling

Mistake: thinking noindex stops the crawling

People slap noindex on thin pages and assume Googlebot will stop visiting them. Then they're shocked the logs still show daily hits.

Actually, that's by design. To see a noindex tag, Google has to crawl the page first. And it keeps re-crawling — it can't know you didn't remove the tag unless it checks. Noindex is an indexing directive, not a crawling one.

If your real goal is to stop the crawl, robots.txt disallow is the lever. But careful — disallow blocks the crawl, so Google never sees the noindex, and a disallowed URL with external links can still get indexed URL-only.

Pick the actual problem you're solving. Index bloat? Noindex. Crawl waste? Disallow. They are not the same knob.
Этот пост опубликован в Telegram-канале Budget Myths. Подписаться можно по ссылке: @CrawlBudgetMyths.
tech

Свежие посты в категории «Tech Infrastructure»

Все каналы категории →

start

Готовы запустить рекламу через сеть public.tg?

Новый оффер, продукт, GEO, кейс, событие или партнёрский запуск — соберём маршрут под задачу и отдадим медиаплан.

Telegram для медиаплана: @AFFtop_connect. Быстрый тест: $20 за канал, $1000 за пакет по сети.