Search Central LiveDeep Dive Europe 2026

Knowledge base v2.13.0 · Community edition · data through 2 October 2026

Thing · Directive

Allow

The robots.txt rule that lets a crawler fetch a path inside a disallowed one; the most specific matching rule wins.

Claims
11
In Google’s docs
1
Said at the event
7
Not in docs
0
Kit items
6

Google’s documentation 1

Documented in

Said at the event 7

Slide and stage claims that name it, the ones Google’s documentation does not cover first.

Consistent with docs 6

StageConsistent with docsD1-C516

Robots.txt was made extremely simple by design so that anyone could implement and understand it; it was extended as websites grew more complex, but it still has only three rules: user-agent (a named crawler, or * for every crawler), disallow and allow.

Day 1 · How Google interprets robots.txt

StageConsistent with docsD1-C519

A user-agent line with its allow and disallow rules forms a user-agent group, and one group can name several crawlers (the talk's example: Googlebot and Bingbot); give named crawlers their own group only when they need rules the * group should not grant to every crawler.

Day 1 · How Google interprets robots.txt

StageConsistent with docsD1-C531

In the talk's quiz, the answer for letting Googlebot crawl /staging/preview/ while keeping the rest of /staging/ closed to it was a user-agent: googlebot group with both rules, disallow /staging/ and allow /staging/preview/, rather than an allow in the * group or a Googlebot group with only the allow.

Day 1 · How Google interprets robots.txt

Confirmed by docs 1

Press and analysis 3

AnalysisD1-C083

Under the example file, Googlebot may crawl /, /politics/eu-vote and /sports/live/, and is blocked from /?utm_source=x, /index.html, /politics (no trailing slash), /live/ and /sports/live-score. An 'allow: /$' rule does not cover the homepage with tracking parameters.

Ibrahim Anjro · Day 1 · How Google interprets robots.txt

AnalysisD1-C532

A typo in a rule name is not always ignored by Google: its open-source robots.txt parser deliberately accepts common misspellings of disallow (such as dissallow, dissalow and disalow) and of user-agent (useragent, user agent), but not of allow. Google's spec page does not mention typos, and other crawlers may be stricter, so spell rule names correctly.

Ibrahim Anjro · Day 1 · How Google interprets robots.txt

Built on these claims 6

Kit items about Allow: their own words name it, or several of the claims they rest on do.

Developer requirements 5

Also inglossary term User-agent group

Connected things 6

Relations

Most often named with it

Things named in the same claim, with the number of claims they share.

Topics that feature it