Search Central LiveDeep Dive Europe 2026

Knowledge base v2.13.0 · Community edition · data through 2 October 2026

Topic · The index and its signals

Indexing statuses in Search Console

Google described the two not-indexed statuses as different stages: 'Discovered – currently not indexed' is a crawl-scheduling state in which Google knows the URL but does not want to crawl it yet, and 'Crawled – currently not indexed' is an index selection decision taken after processing. On stage Google called the first the nastier one and said owners can influence it by getting other URLs indexed and showing that the site's content is useful, while the Page indexing help explains it by expected server overload. For 'Crawled – currently not indexed', Google said the cause is most of the time quality rather than a technical fault: check whether the pages match the quality of the indexed parts of the site, or whether a different template hides where their content is; the help page adds that such pages need no resubmission. Other not-indexed reasons include noindex and 'Alternate page with proper canonical tag', the canonical choice of deduplication that site owners see in Search Console, and Google suggested the report's reasons for checking index selection issues when testing changes. Google's URL Inspection help says a valid live test only confirms that Google can access a page; indexing still depends on other conditions, such as not being a duplicate and being of high enough quality. On Day 3 a community speaker said analysing Search Console query data with an LLM helped his agency spot developer mistakes such as unwanted pages being indexed, which Search Console also shows directly. On Day 1 Gary Illyes had already pointed to the report's breakdown of reasons as a way to find patterns in how a site's content is crawled and served. An audio recording of Day 1 added that soft 404s are looked up in this report rather than in Crawl Stats, and Dave Smart warned in Lightning session B that a URL reported as blocked by robots.txt may not be disallowed itself: Search Console reports a block anywhere in a redirect chain on the chain's first URL, so check whether the URL redirects and test every URL in the chain.

Based on D2-C706, D2-C711, D2-C709, D2-C707, D2-C714, D2-C715, D2-C713, D2-C712, D2-C718, D2-C359, D2-C717, D2-C308, D2-C708, D2-C710, D2-C716, D3-C488, D1-C368, D1-C508, D1-C535, D1-C537

22 claims · raised in 6 sessions · said or shown on Day 1 and Day 2 and Day 3

Open in Reef mapOpen in Graph

What to do

  • For 'Discovered – currently not indexed', rule out slow responses and server errors first, then improve the site's indexed pages instead of resubmitting URLs.
  • For 'Crawled – currently not indexed', compare the URLs with indexed pages of the same type for quality, duplication, soft-404 symptoms and template differences.
  • Use the report's not-indexed reasons to measure the effect of changes to the site.
  • Do not read a passing live test in URL Inspection as a promise of indexing.
  • When a URL is reported as blocked by robots.txt but the file does not block it, test every URL in its redirect chain.

Day 1: Crawling 4

Said on stage 4

StageConfirmed by docsD1-C368

Search Console's page indexing report breaks down the reasons why pages do or do not show in Search, and its categories help find patterns in how a site's content is crawled and served to Google's crawlers.

Speaker Gary IllyesIn Day 1, 14:35 · How crawling errors affect SearchEvidence transcript

  • Extended by D1-C508 Day 1: Google said soft 404s, like some other problem categories, are looked up in Search Console's page indexing…
  • Extended by D2-C717 Day 2: Search Console's Page indexing report is the place to check for index selection issues, and its not-indexed…
StageConfirmed by docsD1-C508

Google said soft 404s, like some other problem categories, are looked up in Search Console's page indexing report, while other crawl issues are debugged in the Crawl Stats report.

Speaker Gary IllyesIn Day 1, 14:35 · How crawling errors affect SearchEvidence transcript

Used byrequirement DEV-MON-03

  • Extends D1-C368 Day 1: Search Console's page indexing report breaks down the reasons why pages do or do not show in Search, and its…
  • Extends D1-C367 Day 1: CDN captcha challenges often return HTTP 200; Googlebot does not solve them and sees only content with a 200…
StageConsistent with docsD1-C535

Dave Smart said Search Console reports such a block only on the first URL of the chain, which is confusing: the URL shown as blocked by robots.txt is not itself disallowed.

Speaker Dave SmartIn Day 1, 15:30 · Lightning session B: Robots.txtEvidence transcript

Used byrequirement DEV-CAN-11glossary term Redirect chain

Day 2: Indexing 17

Said on stage 11

StageConsistent with docsD2-C306

To use Search Console to debug what Google can see on a site and why, Erin Sparling said the first step is to get access to the site by verifying ownership (the speaker's words were authorized domains).

Speaker Erin SparlingIn Day 2, 11:15 · What is Google friendly JavaScriptEvidence transcript

Used byrequirement DEV-MON-01

StageConsistent with docsD2-C359

Google's duplication talk described three related parts of deduplication: building clusters, localization, and selecting the representative URL, which is the canonicalization site owners see in Search Console.

Speaker John MuellerIn Day 2, 11:55 · Handling web duplicationEvidence transcript

StageNot in docsD2-C706

'Discovered – currently not indexed' in Search Console is a crawl-scheduling state: Google knows the URL exists but does not want to crawl it yet. Of the two not-indexed statuses discussed, it was called the 'kind of nastier' one.

“The first one is kind of nastier.”

Speaker GoogleIn Day 2, 15:40 · Deciding what goes in the index?Evidence transcript

Used byrequirement DEV-MON-03glossary term Discovered – currently not indexed

  • Extends D1-C097 Day 1: Google's crawl budget guide is written for sites with over 1 million unique pages that change weekly, over…
  • Extends D1-C065 Day 1: The scheduler is shared infrastructure that decides what to fetch and when and sends URLs to the crawler.…
StageConsistent with docsD2-C709

Site owners can influence 'Discovered – currently not indexed' by getting other URLs of the site indexed and showing Google's systems that the site's content is good and useful to users.

Speaker GoogleIn Day 2, 15:40 · Deciding what goes in the index?Evidence transcript

  • Extends D1-C093 Day 1: Crawl demand is driven by the quality of the site, the change frequency of its URLs and their popularity on…
StageConsistent with docsD2-C711

'Crawled – currently not indexed' in Search Console is an index selection decision: Google crawled and processed the page but decided not to keep it in the index.

Speaker GoogleIn Day 2, 15:40 · Deciding what goes in the index?Evidence transcript

Used byrequirement DEV-MON-03glossary term Crawled – currently not indexed

StageNot in docsD2-C714

'Crawled – currently not indexed' is most of the time a quality issue rather than a technical one: the pages are usually low quality or useless for the index, for example duplicates or soft 404s.

“most of the time it is actually a quality issue”

Speaker GoogleIn Day 2, 15:40 · Deciding what goes in the index?Evidence transcript

Things

Used byrequirement DEV-MON-03

  • Extends D1-C098 Day 1: On smaller sites, slow indexing is almost always a demand problem, meaning quality, not a capacity problem.
StageConsistent with docsD2-C717

Search Console's Page indexing report is the place to check for index selection issues, and its not-indexed reasons are useful when testing changes on a site.

Speaker GoogleIn Day 2, 15:40 · Deciding what goes in the index?Evidence transcript

Used byrequirement DEV-MON-03

  • Extends D1-C368 Day 1: Search Console's page indexing report breaks down the reasons why pages do or do not show in Search, and its…

What Google's documentation says 3

DocsSourceD2-C308

Google's URL Inspection help says a valid live test only confirms that Google can access a page for indexing; the page must still meet other conditions to be indexed, such as having no manual action, not being a duplicate and being of high enough quality.

Publisher Google Search Console HelpAnnotates Day 2, 11:15 · What is Google friendly JavaScript

Used byrequirement DEV-MON-02

DocsSourceD2-C707

Google's Page indexing report help says a 'Discovered – currently not indexed' page was found but not crawled yet, typically because Google wanted to crawl it but expected the crawl to overload the site, so it rescheduled the crawl.

Publisher Google Search Console HelpAnnotates Day 2, 15:40 · Deciding what goes in the index?

Used byrequirement DEV-MON-03glossary term Discovered – currently not indexed

DocsSourceD2-C712

Google's Page indexing report help says a 'Crawled – currently not indexed' page was crawled but not indexed, may or may not be indexed in the future, and does not need to be resubmitted for crawling.

Publisher Google Search Console HelpAnnotates Day 2, 15:40 · Deciding what goes in the index?

Used byrequirement DEV-MON-03glossary term Crawled – currently not indexed

Analysis by the author 3

AnalysisD2-C708

The help page explains 'Discovered – currently not indexed' by expected server overload (capacity), while on stage it was explained as Google not wanting the URL yet (demand); the crawl budget guide covers both, so first rule out slow responses and server errors, then treat the status as a quality and demand problem.

Author Ibrahim AnjroAnnotates Day 2, 15:40 · Deciding what goes in the index?

Used byrequirement DEV-MON-03

  • Extends D1-C098 Day 1: On smaller sites, slow indexing is almost always a demand problem, meaning quality, not a capacity problem.
AnalysisD2-C716

Treat 'Crawled – currently not indexed' as a quality audit list: compare those URLs with indexed pages of the same type for thin, duplicate or soft-404-like content and for template differences before looking for technical faults.

Author Ibrahim AnjroAnnotates Day 2, 15:40 · Deciding what goes in the index?

Used byrequirement DEV-MON-03

  • Extends D1-C098 Day 1: On smaller sites, slow indexing is almost always a demand problem, meaning quality, not a capacity problem.

Day 3: Serving: Ranking, Search Console, and Performance 1

Said on stage 1

Across days and sessions 9

  1. Stage D1-C508 Day 1 · How crawling errors affect Search

    Google said soft 404s, like some other problem categories, are looked up in Search Console's page indexing report, while other crawl issues are debugged in the Crawl Stats report.

    extends
    Stage D1-C367 Day 1 · How crawling errors affect Search

    CDN captcha challenges often return HTTP 200; Googlebot does not solve them and sees only content with a 200 status, which indexing then classifies as a soft 404, reported as an error in Search Console.

  2. Stage D1-C508 Day 1 · How crawling errors affect Search

    Google said soft 404s, like some other problem categories, are looked up in Search Console's page indexing report, while other crawl issues are debugged in the Crawl Stats report.

    extends
    Stage D1-C368 Day 1 · How crawling errors affect Search

    Search Console's page indexing report breaks down the reasons why pages do or do not show in Search, and its categories help find patterns in how a site's content is crawled and served to Google's crawlers.

  3. Stage D2-C706 Day 2 · Deciding what goes in the index?

    'Discovered – currently not indexed' in Search Console is a crawl-scheduling state: Google knows the URL exists but does not want to crawl it yet. Of the two not-indexed statuses discussed, it was called the 'kind of nastier' one.

    extends
    Slide D1-C065 Day 1 · How crawling works

    The scheduler is shared infrastructure that decides what to fetch and when and sends URLs to the crawler. Each team decides the scheduling parameters for its own user agents.

  4. Stage D2-C706 Day 2 · Deciding what goes in the index?

    'Discovered – currently not indexed' in Search Console is a crawl-scheduling state: Google knows the URL exists but does not want to crawl it yet. Of the two not-indexed statuses discussed, it was called the 'kind of nastier' one.

    extends
    Docs D1-C097 Day 1 · How Google thinks about crawl budget

    Google's crawl budget guide is written for sites with over 1 million unique pages that change weekly, over 10,000 pages that change daily, or many URLs reported as 'Discovered – currently not indexed'.

  5. Analysis D2-C708 Day 2 · Deciding what goes in the index?

    The help page explains 'Discovered – currently not indexed' by expected server overload (capacity), while on stage it was explained as Google not wanting the URL yet (demand); the crawl budget guide covers both, so first rule out slow responses and server errors, then treat the status as a quality and demand problem.

    extends
    Analysis D1-C098 Day 1 · How Google thinks about crawl budget

    On smaller sites, slow indexing is almost always a demand problem, meaning quality, not a capacity problem.

  6. Stage D2-C709 Day 2 · Deciding what goes in the index?

    Site owners can influence 'Discovered – currently not indexed' by getting other URLs of the site indexed and showing Google's systems that the site's content is good and useful to users.

    extends
    Slide D1-C093 Day 1 · How Google thinks about crawl budget

    Crawl demand is driven by the quality of the site, the change frequency of its URLs and their popularity on the internet.

  7. Stage D2-C714 Day 2 · Deciding what goes in the index?

    'Crawled – currently not indexed' is most of the time a quality issue rather than a technical one: the pages are usually low quality or useless for the index, for example duplicates or soft 404s.

    extends
    Analysis D1-C098 Day 1 · How Google thinks about crawl budget

    On smaller sites, slow indexing is almost always a demand problem, meaning quality, not a capacity problem.

  8. Analysis D2-C716 Day 2 · Deciding what goes in the index?

    Treat 'Crawled – currently not indexed' as a quality audit list: compare those URLs with indexed pages of the same type for thin, duplicate or soft-404-like content and for template differences before looking for technical faults.

    extends
    Analysis D1-C098 Day 1 · How Google thinks about crawl budget

    On smaller sites, slow indexing is almost always a demand problem, meaning quality, not a capacity problem.

  9. Stage D2-C717 Day 2 · Deciding what goes in the index?

    Search Console's Page indexing report is the place to check for index selection issues, and its not-indexed reasons are useful when testing changes on a site.

    extends
    Stage D1-C368 Day 1 · How crawling errors affect Search

    Search Console's page indexing report breaks down the reasons why pages do or do not show in Search, and its categories help find patterns in how a site's content is crawled and served to Google's crawlers.

Built on these claims 5

Developer requirements 5

Sources 6