- Class
- C1 Mechanism
- Weight
- 8
- Severity on fail
- critical
- Applies to
- Google and Bing
Test this rule on a live site with the robots.txt checker, the XML sitemap checker, the noindex checker or the AI crawler checker. Free, no account needed.
Scope and evidence
Evaluated on each member of I. The book marks this requirement ELIGIBILITY REQUIREMENT, SEARCH-ENGINE-SPECIFIC, and the scanner treats it as a gate: indexing rule: a failure on an intended page removes that page from the eligible set instead of entering the achieved share.
Fail conditions
Each condition has a stable code and a reason template. The evaluator fills the placeholders from observed values only; no sentence in a report is generated.
disallowed_googlebot{url} is disallowed for Googlebot by {rule} (robots.txt line {line}, group {group}). That Googlebot isn't blocked is one of Google's technical requirements for indexing. Google may still index the URL from links elsewhere, but without reading the page, and the result won't have a description.
disallowed_bingbot{url} is disallowed for bingbot by {rule} (robots.txt line {line}, group {group}). Under the Robots Exclusion Protocol, bingbot should not crawl it.
The documented fix
Remove or narrow the rule. If the URL should not be in search results, take it out of sitemaps and use noindex on a crawlable URL instead, because robots.txt does not keep pages out of Google.
Sources, quoted
These are the passages the rule rests on, stored once in the knowledge base and emitted exactly as written. A citation marked for specific conditions appears only on findings with those codes.
Cited for context; no quote stored.
[G-TECH] Google Search technical requirements · Googlebot isn't blocked · last updated 2025-12-18 · for disallowed_googlebot
https://developers.google.com/search/docs/essentials/technical“can still be indexed if linked to from other sites”
[G-ROBOTS-INTRO] Introduction to robots.txt · Understand the limitations of a robots.txt file · last updated 2025-12-10 · for disallowed_googlebot
https://developers.google.com/search/docs/crawling-indexing/robots/intro“the search result won't have a description”
[G-ROBOTS-INTRO] Introduction to robots.txt · What is a robots.txt file used for? · last updated 2025-12-10 · for disallowed_googlebot
https://developers.google.com/search/docs/crawling-indexing/robots/intro“These lines indicate whether accessing a URI that matches the corresponding path is allowed or disallowed.”
[RFC9309] RFC 9309: Robots Exclusion Protocol · 2.2.2 The "Allow" and "Disallow" Lines · last updated September 2022 · for disallowed_bingbot
https://www.rfc-editor.org/rfc/rfc9309.html“BingBot will ignore all the other directives”
[B-ROBOTS-2012] To crawl or not to crawl, that is BingBot's question · robots.txt · last updated posted 2012-05-03 · for disallowed_bingbot
https://blogs.bing.com/webmaster/May-2012/To-crawl-or-not-to-crawl,-that-is-BingBot-s-questi