Search Central LiveDeep Dive Europe 2026

Knowledge base v2.13.0 · Community edition · data through 2 October 2026

Thing · Metric

Crawl budget

The number of URLs Google can and wants to crawl on a site, set by its crawl capacity limit and crawl demand.

Metric

Open in Reef mapOpen in Graph

Narrative see the topic Crawl budget

Claims
39
In Google’s docs
4
Said at the event
27
Not in docs
0
Kit items
10

Glossary · Crawl budget

How much Google crawls a site: the combination of crawl rate limit and crawl demand, the finite resources Google allocates to crawling one website. A site with fewer than a few thousand URLs likely has no crawl budget problem, Google said.

Google’s documentation 4

Documented in

DocsSourceD1-C097

Google's crawl budget guide is written for sites with over 1 million unique pages that change weekly, over 10,000 pages that change daily, or many URLs reported as 'Discovered – currently not indexed'.

Google · Day 1 · How Google thinks about crawl budget

Said at the event 27

Slide and stage claims that name it, the ones Google’s documentation does not cover first.

Consistent with docs 16

StageConsistent with docsD1-C373

Crawl budget is the finite amount of resources Google allocates to crawling a specific website, and it determines how many of the site's pages are discovered and how often they are revisited.

Day 1 · How Google thinks about crawl budget

StageConsistent with docsD1-C378

Google treats URLs that differ as different URLs and crawls all of them even when they lead to the same content, so infinite URL spaces such as calendars or many versions of a page burn crawl budget.

Day 1 · How Google thinks about crawl budget

StageConsistent with docsD1-C445

One fix for parameter variants of a URL is to work out the normalised URL and redirect the variants to it: the redirects cost crawl budget at first, but leave the site with a clean slate.

Day 1 · Q&A

StageConsistent with docsD1-C466

Gary Illyes said news sites rarely need to worry about crawl budget, because Google crawls them aggressively given their constant flow of new content and new URLs.

Gary Illyes · Day 1 · Q&A

StageConsistent with docsD1-C469

Gary Illyes said disallowing a section that should not be crawled, such as /ads, in a Googlebot group in robots.txt shifts crawl budget to the rest of the site.

Gary Illyes · Day 1 · Q&A

StageConsistent with docsD1-C539

A Google panelist said useless plugin-generated parameter URLs can be handled at the web server, for example with a rule or a custom module in Apache, which saves crawl budget so that Google may pick up the URLs that matter instead (the exact mechanism is not clear in the recording).

Day 1 · Q&A

StageConsistent with docsD2-C848

Google said it strongly believes site owners should be able to opt out of crawling and control how their site is crawled, which can matter for legal reasons or for crawl budget.

Day 2 · Welcome to indexing day!

Confirmed by docs 6

StageConfirmed by docsD1-C381

Not every site needs to worry about crawl budget, and this has always been so: a site with fewer than a few thousand URLs likely has no crawl budget problem, and a larger site does not necessarily have one either.

Day 1 · How Google thinks about crawl budget

Nothing to verify 5

StageD1-C441

An audience member asked how to tell, from log files or by other methods, whether Google is spending crawl budget inefficiently on a site.

From the audience · Day 1 · Q&A

StageD1-C465

An audience member asked how a large news site can tell whether crawl budget is limiting how fast new articles are discovered (within minutes), and which statistics in Search Console and the logs show this.

From the audience · Day 1 · Q&A

StageD1-C545

An audience member asked, for very large sites of about 100 million pages, what signs show that a site is limited by crawl budget, and how to tell a crawl capacity limit problem from a crawl demand problem.

From the audience · Day 1 · Q&A

Press and analysis 8

AnalysisD1-C382

The stage rule of thumb (under a few thousand URLs crawl budget is unlikely to be a problem) and Google's crawl budget guide (D1-C097: sites with over a million pages that change about weekly, or over 10,000 pages that change daily) leave a middle range where a site should check Search Console's Crawl Stats report before blaming crawl budget for slow indexing.

Ibrahim Anjro · Day 1 · How Google thinks about crawl budget

AnalysisD1-C386

The stage point that 4xx responses do not affect crawl budget matches Google's documentation that 4xx codes have no effect on crawl rate (D1-C072), with one exception: 429 Too Many Requests counts as a server error and slows crawling like a 5xx (D1-C071, D1-C092).

Ibrahim Anjro · Day 1 · How Google thinks about crawl budget

AnalysisD1-C419

A 404 fetch is still a fetch: Google's 2017 crawl budget post says generally any URL Googlebot crawls counts towards a site's crawl budget, so the stage point that 4xx responses do not affect crawl budget is best read as 'they do not slow crawling, and a 404 tells Google to crawl that URL less over time'.

Ibrahim Anjro · Day 1 · How Google thinks about crawl budget

AnalysisD1-C420

The stage remark that crawl demand follows the quality of the site as a whole sits beside the same talk's slide (D1-C094), which falls back to the parent path's aggregate only when a URL's own quality is unknown, and Google's crawl budget guide lists page quality among the demand factors; read it as site quality setting the baseline while known URL-level signals still count.

Ibrahim Anjro · Day 1 · How Google thinks about crawl budget

AnalysisD1-C470

Google's crawl budget guide says crawl budget freed by robots.txt blocks is not shifted to other pages unless the site already hits its crawl capacity limit, and advises against robots.txt for temporary reallocation; so block only sections you never want crawled, and expect a shift only on capacity-limited sites.

Ibrahim Anjro · Day 1 · Q&A

AnalysisD1-C547

Google's crawl budget guide does not express the crawl capacity limit in requests per second: it limits the total time a server spends holding connections open for Google, counting both the number of parallel connections and their duration. That fits the panel's advice to watch how many connections Googlebot opens (D1-C494): fewer connections is the documented form of a lower capacity limit.

Ibrahim Anjro · Day 1 · Q&A

AnalysisD2-C914

Google's crawl budget guide is written for sites with over a million pages changing weekly or over 10,000 pages changing daily, so for a consolidation of about 2,000 URLs the crawl-budget gain is likely minor; the benefit of pruning such a site more plausibly comes from one strong URL per intent and consolidated signals.

Ibrahim Anjro · Day 2 · Lightning session E: Managing Duplicates and Site Moves

AnalysisD2-C708

The help page explains 'Discovered – currently not indexed' by expected server overload (capacity), while on stage it was explained as Google not wanting the URL yet (demand); the crawl budget guide covers both, so first rule out slow responses and server errors, then treat the status as a quality and demand problem.

Ibrahim Anjro · Day 2 · Deciding what goes in the index?

Built on these claims 10

Kit items about Crawl budget: their own words name it, or several of the claims they rest on do.

Developer requirements 8

3 more

Facts 1

Also inglossary term Crawl budget

Connected things 20

Relations

  • Has part Crawl capacity limit structure, no claim needed
  • Has part Crawl demand structure, no claim needed
  • Affected by Faceted navigation
    1 claim, 1 documented
    • SlideConsistent with docsD1-C103

      Four ways to manage crawl budget: use HTTP cache control, have good site navigation, restrict crawlers' access to faceted navigation and action URLs, and improve or remove useless content.

  • Measured by Crawl stats
    1 claim, 1 documented
    • StageConfirmed by docsD1-C393

      To check whether a site has a crawl budget problem, use Search Console's crawl report (Crawl Stats), which breaks crawl requests down by response and by file type and shows crawl problems Google finds.

Most often named with it

Things named in the same claim, with the number of claims they share.

Topics that feature it