Search Central LiveDeep Dive Europe 2026

Knowledge base v2.13.0 · Community edition · data through 2 October 2026

Topic · Crawling

Crawl rate limit (hostload)

Hostload, the crawl capacity limit, is a host-wide limit shared by all Google crawlers, so heavy crawling by one Google product can leave less capacity for the others. It falls when connect time or time to first byte rise, or when the server returns 429, 5xx or network errors. On Day 2 a Google slide showed an example fetch record, as passed on to processing, that includes connect time and time to first byte in milliseconds (not in Google's docs). Day 3 gave timings: when a site starts serving 500 errors, Google lowers its crawl capacity within about four hours on average, while increases take one to three weeks because Google first needs to know that higher demand will last, and a continuously running process recalculates capacity within about a month (the spoken and slide figures differ slightly and are not in Google's docs). Google's crawl rate guide confirms that many 500, 503 or 429 responses reduce the crawl rate, which recovers automatically once the errors drop. The second recording of Day 1 added that crawl rate limit, or hostload, is a proxy for how many requests per second the connection to a site can handle, and that it applies per host, not necessarily per site: a site's CDN, app, subdomains and www host may be different hosts. Cherry Prommawin said 429 is the one 4xx code that tells Google to slow down, and that a 5xx makes Google slow down so as not to break the site; a rise in errors can come from a CDN throttling crawlers by injecting 429 or 503 responses on the network path. In the Q&A Google said that when higher demand brings more crawling than a server can cope with, Google reduces it again, and, asked how a site of about 100 million pages can tell a crawl capacity problem from a crawl demand problem, that a capacity drop is most of the time an abrupt step down, illustrated as hostload falling from 10 to 5, which a panelist said means up to five requests per second, and that a fall in the number of connections Googlebot opens shows the capacity limit changed. Google added that demand from Search, Google Ads or another Google product raises crawling only as far as the site's crawl capacity allows. Author’s view: Google's crawl budget guide does not express the capacity limit in requests per second but as the total time a server spends holding connections open for Google, counting parallel connections and their duration, so fewer connections is the documented form of a lower limit. Author’s view: to slow crawling for a short time return 429 or 503, never 401, 402 or 403, which Google treats like 404.

What to do

  • Keep server response times stable; a slow server lowers crawling for every Google product on that host.
  • Return 429 or 5xx only for real overload or outages; they lower hostload for every Google crawler on the host.
  • Keep servers stable before launches and migrations: a burst of 5xx errors cuts crawling within hours and recovery takes weeks.
  • Check the number of connections Googlebot opens to your server to see whether Google lowered your crawl capacity limit.
  • Treat each subdomain, CDN host and app host as its own hostload when planning server capacity.

Day 1: Crawling 22

Shown on screen 3

SlideConsistent with docsD1-C066

Google Search, Ads, Shopping and Images all request through one centralised crawling infrastructure, whose primary mandate is to fetch from the internet while strictly preventing the overloading of external servers.

Speaker Cherry PrommawinIn Day 1, 16:00 · How Google thinks about crawl budgetEvidence slide photo, transcript

Used byrequirement DEV-PRF-01

  • Extends D1-C326 Day 1: Google does not let each team build its own crawler: its many crawlers share one crawler infrastructure…
  • Extended by D1-C375 Day 1: Googlebot as a single standalone crawler is a historical idea: Google crawls through a centralised crawling…
SlideConfirmed by docsD1-C091

Crawl rate limit, or hostload, is a host-wide metric shared across all Google crawlers.

Speaker Cherry PrommawinIn Day 1, 16:00 · How Google thinks about crawl budgetEvidence slide photo, transcript

Used byrequirement DEV-PRF-01glossary term Crawl rate limit (hostload)

  • Extended by D1-C374 Day 1: Crawl rate limit, or hostload, is a proxy for how many requests per second the connection to a site can…
  • Extended by D1-C376 Day 1: Hostload applies per host, not necessarily per site: depending on how a site is configured, its CDN, its app…
SlideConsistent with docsD1-C092

Hostload is driven by changes in connect time, changes in time to first byte, and HTTP 429 or 5xx status codes. If these increase, hostload is adjusted and crawling slows down.

Speaker Cherry PrommawinIn Day 1, 16:00 · How Google thinks about crawl budgetEvidence slide photo, transcript

Used byrequirements DEV-MON-04, DEV-PRF-01glossary term Crawl rate limit (hostload)

  • Extended by D2-C024 Day 2: A Google slide titled Data from Crawling showed an example of what the crawler passes on for processing: the…

Said on stage 12

StageConfirmed by docsD1-C354

Google slows crawling when a site returns 5xx errors, because a 5xx usually means the server, and often the whole site, cannot serve requests, and Google does not want to break the site.

Speaker Cherry PrommawinIn Day 1, 14:35 · How crawling errors affect SearchEvidence transcript

Things

Used byrequirement DEV-SRV-03

  • Extended by D3-C618 Day 3: When a site starts serving 500 errors, Google lowers the crawl capacity allocated to the site within about…
StageConsistent with docsD1-C374

Crawl rate limit, or hostload, is a proxy for how many requests per second the connection to a site can handle before Google hits the limit.

Speaker Cherry PrommawinIn Day 1, 16:00 · How Google thinks about crawl budgetEvidence transcript

Used byrequirement DEV-PRF-01

  • Extends D1-C091 Day 1: Crawl rate limit, or hostload, is a host-wide metric shared across all Google crawlers.
  • Extended by D1-C546 Day 1: A Google panelist illustrated a crawl capacity limit drop as a step down in the site's hostload, for example…
StageConfirmed by docsD1-C376

Hostload applies per host, not necessarily per site: depending on how a site is configured, its CDN, its app, its subdomains and its main www host may be different hosts.

Speaker Cherry PrommawinIn Day 1, 16:00 · How Google thinks about crawl budgetEvidence transcript

Used byrequirement DEV-PRF-01glossary term Crawl rate limit (hostload)

  • Extends D1-C091 Day 1: Crawl rate limit, or hostload, is a host-wide metric shared across all Google crawlers.
StageConsistent with docsD1-C439

Very large, constantly changing sites get no special handling: when URLs change frequently or are useful to users, Google raises crawl demand and tries to raise its crawl capacity for the site, the same logic it applies to small sites.

Speaker not identifiedIn Day 1, 16:35 · Q&AEvidence transcript

  • Answers D1-C438 Day 1: An audience member asked how Google prioritises crawling and indexing for very large real-time sites, such as…
  • Extends D1-C093 Day 1: Crawl demand is driven by the quality of the site, the change frequency of its URLs and their popularity on…
  • Extended by D1-C538 Day 1: A Google panelist said crawl demand is not only Search's: when enough of a site's URLs are wanted by Search…
  • Extended by D3-C623 Day 3: When Google notices the web getting excited about a few URLs on a site, it allocates more crawl demand to the…
StageConfirmed by docsD1-C440

If a site's server cannot cope with the extra crawling that higher crawl demand brings, Google reduces its crawling of that site again.

Speaker not identifiedIn Day 1, 16:35 · Q&AEvidence transcript

Used byrequirement DEV-PRF-01

  • Answers D1-C438 Day 1: An audience member asked how Google prioritises crawling and indexing for very large real-time sites, such as…
  • Extended by D3-C618 Day 3: When a site starts serving 500 errors, Google lowers the crawl capacity allocated to the site within about…
StageD1-C545

An audience member asked, for very large sites of about 100 million pages, what signs show that a site is limited by crawl budget, and how to tell a crawl capacity limit problem from a crawl demand problem.

From the audienceIn Day 1, 16:35 · Q&AEvidence transcript

  • Answered by D1-C493 Day 1: There is no easy way to tell whether a drop in Google's crawling comes from lower crawl demand or a lower…
  • Answered by D1-C494 Day 1: To check whether Google lowered a site's crawl capacity limit, look at the number of connections Googlebot…
  • Answered by D1-C546 Day 1: A Google panelist illustrated a crawl capacity limit drop as a step down in the site's hostload, for example…
StageConsistent with docsD1-C493

There is no easy way to tell whether a drop in Google's crawling comes from lower crawl demand or a lower crawl capacity limit; most of the time a capacity-limit drop is an abrupt step down.

Speaker not identifiedIn Day 1, 16:35 · Q&AEvidence transcript

Used byrequirement DEV-PRF-01glossary term Crawl rate limit (hostload)

  • Answers D1-C545 Day 1: An audience member asked, for very large sites of about 100 million pages, what signs show that a site is…
  • Extended by D3-C617 Day 3: Google's crawl chart puts a crawl capacity update at seconds when backing off, typically 4 hours or 1-2…
StageConsistent with docsD1-C494

To check whether Google lowered a site's crawl capacity limit, look at the number of connections Googlebot opens to the site: if it has dropped, the capacity limit changed.

Speaker not identifiedIn Day 1, 16:35 · Q&AEvidence transcript

Used byrequirement DEV-PRF-01glossary term Crawl rate limit (hostload)

  • Answers D1-C545 Day 1: An audience member asked, for very large sites of about 100 million pages, what signs show that a site is…
StageConsistent with docsD1-C538

A Google panelist said crawl demand is not only Search's: when enough of a site's URLs are wanted by Search, Google Ads or another Google product, Google raises the site's crawl demand and crawls up to that level of demand as far as the site's crawl capacity allows.

Speaker not identifiedIn Day 1, 16:35 · Q&AEvidence transcript

Used byglossary term Crawl demand

  • Extends D1-C439 Day 1: Very large, constantly changing sites get no special handling: when URLs change frequently or are useful to…
  • Extends D1-C334 Day 1: Different Google teams prioritise crawling differently: web search cares a lot about the quality of a site…
  • Repeated by D3-C624 Day 3: Crawl demand can also come from other Google products, such as Shopping, and the roughly 20-hour estimate…
StageConsistent with docsD1-C546

A Google panelist illustrated a crawl capacity limit drop as a step down in the site's hostload, for example from 10 to 5, which the panelist said means Googlebot then makes up to five requests per second (the recording adds 'per connection', which is unclear).

Speaker not identifiedIn Day 1, 16:35 · Q&AEvidence transcript

  • Answers D1-C545 Day 1: An audience member asked, for very large sites of about 100 million pages, what signs show that a site is…
  • Extends D1-C374 Day 1: Crawl rate limit, or hostload, is a proxy for how many requests per second the connection to a site can…
  • Extended by D1-C547 Day 1: Google's crawl budget guide does not express the crawl capacity limit in requests per second: it limits the…

What Google's documentation says 3

DocsSourceD1-C126

Google treats network timeouts, connection resets and DNS errors like 5xx server errors: crawling slows down immediately, and already indexed URLs that stay unreachable are removed from Google's index within days.

“Google treats network timeouts, connection reset, and DNS errors similarly to 5xx server errors.”

Publisher GoogleAnnotates Day 1, 14:35 · How crawling errors affect Search

Things

Used byrequirements DEV-MON-04, DEV-SRV-01, DEV-SRV-03fact F-017

  • Extended by D3-C618 Day 3: When a site starts serving 500 errors, Google lowers the crawl capacity allocated to the site within about…
DocsSourceD1-C132

Google's crawl budget guide says each crawler has its own crawl demand, but the crawl capacity limit (hostload) is shared across all crawlers, so high demand from one crawler can reduce the capacity left for others.

“high demand from one crawler can reduce the capacity available for others”

Publisher GoogleAnnotates Day 1, 16:00 · How Google thinks about crawl budget

Used byrequirement DEV-PRF-01glossary term Crawl rate limit (hostload)

Analysis by the author 4

AnalysisD1-C547

Google's crawl budget guide does not express the crawl capacity limit in requests per second: it limits the total time a server spends holding connections open for Google, counting both the number of parallel connections and their duration. That fits the panel's advice to watch how many connections Googlebot opens (D1-C494): fewer connections is the documented form of a lower capacity limit.

Author Ibrahim AnjroAnnotates Day 1, 16:35 · Q&A

Used byrequirement DEV-PRF-01

  • Extends D1-C546 Day 1: A Google panelist illustrated a crawl capacity limit drop as a step down in the site's hostload, for example…

Day 2: Indexing 1

Shown on screen 1

SlideNot in docsD2-C024

A Google slide titled Data from Crawling showed an example of what the crawler passes on for processing: the fetch result, the connect time and the time to first byte in milliseconds, the robots policies that apply to the fetch, and the raw HTTP response with the page's HTML.

Speaker Cherry PrommawinIn Day 2, 10:25 · How is HTML interpretedEvidence slide photo

Used byrequirement DEV-PRF-01

  • Extends D1-C063 Day 1: The crawl pipeline runs from a crawl queue to a scheduler to the crawler, which fetches from the internet and…
  • Extends D1-C064 Day 1: The crawler has multiple tasks: fetch from the internet, ensure it doesn't break the internet, and enforce…
  • Extends D1-C092 Day 1: Hostload is driven by changes in connect time, changes in time to first byte, and HTTP 429 or 5xx status…

Day 3: Serving: Ranking, Search Console, and Performance 6

Shown on screen 1

Said on stage 3

StageConsistent with docsD3-C618

When a site starts serving 500 errors, Google lowers the crawl capacity allocated to the site within about four hours on average.

Speaker Gary IllyesIn Day 3, 15:45 · How long does it take to..?Evidence transcript

Used byrequirement DEV-SRV-03

  • Extends D1-C126 Day 1: Google treats network timeouts, connection resets and DNS errors like 5xx server errors: crawling slows down…
  • Extends D1-C354 Day 1: Google slows crawling when a site returns 5xx errors, because a 5xx usually means the server, and often the…
  • Extends D1-C440 Day 1: If a site's server cannot cope with the extra crawling that higher crawl demand brings, Google reduces its…

What Google's documentation says 1

Analysis by the author 1

Across days and sessions 18

  1. Slide D1-C066 Day 1 · How Google thinks about crawl budget

    Google Search, Ads, Shopping and Images all request through one centralised crawling infrastructure, whose primary mandate is to fetch from the internet while strictly preventing the overloading of external servers.

    extends
    Stage D1-C326 Day 1 · How crawling works

    Google does not let each team build its own crawler: its many crawlers share one crawler infrastructure, because every crawler must accomplish a few specific tasks and obey Google's internal crawling policies.

  2. Stage D1-C374 Day 1 · How Google thinks about crawl budget

    Crawl rate limit, or hostload, is a proxy for how many requests per second the connection to a site can handle before Google hits the limit.

    extends
    Slide D1-C091 Day 1 · How Google thinks about crawl budget

    Crawl rate limit, or hostload, is a host-wide metric shared across all Google crawlers.

  3. Stage D1-C375 Day 1 · How Google thinks about crawl budget

    Googlebot as a single standalone crawler is a historical idea: Google crawls through a centralised crawling infrastructure, so a request from a Google user agent in server logs is a request routed through that shared platform.

    extends
    Slide D1-C066 Day 1 · How Google thinks about crawl budget

    Google Search, Ads, Shopping and Images all request through one centralised crawling infrastructure, whose primary mandate is to fetch from the internet while strictly preventing the overloading of external servers.

  4. Stage D1-C376 Day 1 · How Google thinks about crawl budget

    Hostload applies per host, not necessarily per site: depending on how a site is configured, its CDN, its app, its subdomains and its main www host may be different hosts.

    extends
    Slide D1-C091 Day 1 · How Google thinks about crawl budget

    Crawl rate limit, or hostload, is a host-wide metric shared across all Google crawlers.

  5. Stage D1-C439 Day 1 · Q&A

    Very large, constantly changing sites get no special handling: when URLs change frequently or are useful to users, Google raises crawl demand and tries to raise its crawl capacity for the site, the same logic it applies to small sites.

    extends
    Slide D1-C093 Day 1 · How Google thinks about crawl budget

    Crawl demand is driven by the quality of the site, the change frequency of its URLs and their popularity on the internet.

  6. Stage D1-C538 Day 1 · Q&A

    A Google panelist said crawl demand is not only Search's: when enough of a site's URLs are wanted by Search, Google Ads or another Google product, Google raises the site's crawl demand and crawls up to that level of demand as far as the site's crawl capacity allows.

    extends
    Stage D1-C334 Day 1 · How crawling works

    Different Google teams prioritise crawling differently: web search cares a lot about the quality of a site and its content, while Ads wants to check every publisher page that wants to appear in Google Ads, so it schedules those URLs as they come in.

  7. Stage D1-C538 Day 1 · Q&A

    A Google panelist said crawl demand is not only Search's: when enough of a site's URLs are wanted by Search, Google Ads or another Google product, Google raises the site's crawl demand and crawls up to that level of demand as far as the site's crawl capacity allows.

    extends
    Stage D1-C439 Day 1 · Q&A

    Very large, constantly changing sites get no special handling: when URLs change frequently or are useful to users, Google raises crawl demand and tries to raise its crawl capacity for the site, the same logic it applies to small sites.

  8. Stage D1-C546 Day 1 · Q&A

    A Google panelist illustrated a crawl capacity limit drop as a step down in the site's hostload, for example from 10 to 5, which the panelist said means Googlebot then makes up to five requests per second (the recording adds 'per connection', which is unclear).

    extends
    Stage D1-C374 Day 1 · How Google thinks about crawl budget

    Crawl rate limit, or hostload, is a proxy for how many requests per second the connection to a site can handle before Google hits the limit.

  9. Analysis D1-C547 Day 1 · Q&A

    Google's crawl budget guide does not express the crawl capacity limit in requests per second: it limits the total time a server spends holding connections open for Google, counting both the number of parallel connections and their duration. That fits the panel's advice to watch how many connections Googlebot opens (D1-C494): fewer connections is the documented form of a lower capacity limit.

    extends
    Stage D1-C546 Day 1 · Q&A

    A Google panelist illustrated a crawl capacity limit drop as a step down in the site's hostload, for example from 10 to 5, which the panelist said means Googlebot then makes up to five requests per second (the recording adds 'per connection', which is unclear).

  10. Slide D2-C024 Day 2 · How is HTML interpreted

    A Google slide titled Data from Crawling showed an example of what the crawler passes on for processing: the fetch result, the connect time and the time to first byte in milliseconds, the robots policies that apply to the fetch, and the raw HTTP response with the page's HTML.

    extends
    Slide D1-C063 Day 1 · How crawling works

    The crawl pipeline runs from a crawl queue to a scheduler to the crawler, which fetches from the internet and passes the fetch reply to indexing.

  11. Slide D2-C024 Day 2 · How is HTML interpreted

    A Google slide titled Data from Crawling showed an example of what the crawler passes on for processing: the fetch result, the connect time and the time to first byte in milliseconds, the robots policies that apply to the fetch, and the raw HTTP response with the page's HTML.

    extends
    Slide D1-C064 Day 1 · How crawling works

    The crawler has multiple tasks: fetch from the internet, ensure it doesn't break the internet, and enforce robots.txt policies. Fetched data is sent for indexing.

  12. Slide D2-C024 Day 2 · How is HTML interpreted

    A Google slide titled Data from Crawling showed an example of what the crawler passes on for processing: the fetch result, the connect time and the time to first byte in milliseconds, the robots policies that apply to the fetch, and the raw HTTP response with the page's HTML.

    extends
    Slide D1-C092 Day 1 · How Google thinks about crawl budget

    Hostload is driven by changes in connect time, changes in time to first byte, and HTTP 429 or 5xx status codes. If these increase, hostload is adjusted and crawling slows down.

  13. Slide D3-C617 Day 3 · How long does it take to..?

    Google's crawl chart puts a crawl capacity update at seconds when backing off, typically 4 hours or 1-2 weeks, and 1-3 weeks in recovery.

    extends
    Stage D1-C493 Day 1 · Q&A

    There is no easy way to tell whether a drop in Google's crawling comes from lower crawl demand or a lower crawl capacity limit; most of the time a capacity-limit drop is an abrupt step down.

  14. Stage D3-C618 Day 3 · How long does it take to..?

    When a site starts serving 500 errors, Google lowers the crawl capacity allocated to the site within about four hours on average.

    extends
    Docs D1-C126 Day 1 · How crawling errors affect Search

    Google treats network timeouts, connection resets and DNS errors like 5xx server errors: crawling slows down immediately, and already indexed URLs that stay unreachable are removed from Google's index within days.

  15. Stage D3-C618 Day 3 · How long does it take to..?

    When a site starts serving 500 errors, Google lowers the crawl capacity allocated to the site within about four hours on average.

    extends
    Stage D1-C354 Day 1 · How crawling errors affect Search

    Google slows crawling when a site returns 5xx errors, because a 5xx usually means the server, and often the whole site, cannot serve requests, and Google does not want to break the site.

  16. Stage D3-C618 Day 3 · How long does it take to..?

    When a site starts serving 500 errors, Google lowers the crawl capacity allocated to the site within about four hours on average.

    extends
    Stage D1-C440 Day 1 · Q&A

    If a site's server cannot cope with the extra crawling that higher crawl demand brings, Google reduces its crawling of that site again.

  17. Stage D3-C623 Day 3 · How long does it take to..?

    When Google notices the web getting excited about a few URLs on a site, it allocates more crawl demand to the whole site so it does not miss content useful to future searchers.

    extends
    Stage D1-C439 Day 1 · Q&A

    Very large, constantly changing sites get no special handling: when URLs change frequently or are useful to users, Google raises crawl demand and tries to raise its crawl capacity for the site, the same logic it applies to small sites.

  18. Stage D3-C624 Day 3 · How long does it take to..?

    Crawl demand can also come from other Google products, such as Shopping, and the roughly 20-hour estimate covers only demand from Search.

    repeats
    Stage D1-C538 Day 1 · Q&A

    A Google panelist said crawl demand is not only Search's: when enough of a site's URLs are wanted by Search, Google Ads or another Google product, Google raises the site's crawl demand and crawls up to that level of demand as far as the site's crawl capacity allows.

Built on these claims 6

Developer requirements 5

Facts 1

Sources 9