Robots.txt Generator
Create a clean robots.txt file with sitemap references, safe WordPress defaults and optional rules for AI crawlers, then…
Paste or load a robots.txt file and test which URLs Googlebot, Bingbot or any crawler may crawl, with the exact rule that decides each one.
The tester follows the Robots Exclusion Protocol (RFC 9309), the rules Google and Bing use: the crawler obeys the most specific matching user-agent group, the longest matching Allow or Disallow rule wins, a tie goes to Allow, * matches any characters and $ anchors the end of the URL. It also lists syntax problems such as rules outside a group and unknown directives.
One stray Disallow line can stop search engines from crawling a whole site. Testing important URLs before you change robots.txt prevents expensive mistakes.
A missing file (404) means everything may be crawled. A server error (5xx) can make Google pause crawling the site.
No. Google ignores Crawl-delay; Bing and some other crawlers respect it.
Create a clean robots.txt file with sitemap references, safe WordPress defaults and optional rules for AI crawlers, then…
Find out whether a page tells search engines not to index it, using both the meta robots tag…
Extract every URL from an XML sitemap or sitemap index, including gzipped sitemaps, and export the list to…
Find out which of up to 50 URLs are blocked from search results by noindex in the meta…
Count the visible words on any web page and see its most frequent terms, without copying and pasting…
View the HTTP response headers of any URL and check caching, compression and common security headers such as…
Trace every redirect from a URL to its final destination, with the status code (301, 302, 307, 308)…
Check the HTTP status code a URL returns (200, 301, 404, 500…), how long it took and where…