Crawlability vs Indexability
Separates bot fetch permission (robots.txt) from indexing permission (meta noindex).
robots.txt Rule Audit
Evaluates crawler-specific `Disallow:` and `Allow:` directives with wildcard matching.
Meta Robots & X-Robots-Tag
Scans HTML header tags and HTTP response headers for noindex and nofollow directives.
Auditing Webpage Crawlability & Indexability…
Fetching host robots.txt rules, tracing HTTP redirect chains, and analyzing meta directives.
Crawlability Access
Indexability Status
HTTP Status Code
200
Bot Tested
Googlebot
Crawlability & Indexability Decision Tree Evidence
Crawlability Control (Bot Payload Access)
Determines whether search crawlers can fetch the webpage file from the server.
Indexability Control (Search Results Eligibility)
Determines whether search engines are allowed to index this page in search results.
Host robots.txt Directive Evaluation
NoneDiscovered Sitemaps in robots.txt
- No Sitemap directives discovered in robots.txt
HTML Robots Meta & HTTP X-Robots-Tag Audit
index, followRedirect Chain & Canonical Link Tag
HTTP Redirect Chain
- No redirects. Target loaded directly.
Canonical Link Tag Evaluation
NonePage Links & Crawl Budget Metrics
0
0
0
What is Website Crawlability?
Crawlability describes a search engine crawler's ability to access and scan content pages on a website. If a page is blocked via robots.txt, 500 server errors, or noindex tags, search bots will fail to crawl and index it.