Skip to content

Search WPDesignVault

Try "image", "json", "schema" or "contrast". Use arrow keys to move through results.

SEO Tools

Robots.txt Tester: Check if a URL Is Blocked

Paste or load a robots.txt file and test which URLs Googlebot, Bingbot or any crawler may crawl, with the exact rule that decides each one.

Runs in your browser. Nothing you enter is uploaded. Last updated
Or load from a live URL

Our server fetches the public page and returns only the data this tool needs. Pasted text stays in your browser.

    Test results
    URLResultMatched rule

      How to use

      1. Paste a robots.txt file, or enter a site address and load its live file.
      2. Choose a crawler (user agent) and enter one or more paths or URLs, one per line.
      3. Each line shows Allowed or Blocked and the rule that matched.

      What this tool does

      The tester follows the Robots Exclusion Protocol (RFC 9309), the rules Google and Bing use: the crawler obeys the most specific matching user-agent group, the longest matching Allow or Disallow rule wins, a tie goes to Allow, * matches any characters and $ anchors the end of the URL. It also lists syntax problems such as rules outside a group and unknown directives.

      Why it matters

      One stray Disallow line can stop search engines from crawling a whole site. Testing important URLs before you change robots.txt prevents expensive mistakes.

      Common uses

      • Checking that CSS, JavaScript and images used by pages are not blocked.
      • Testing a new robots.txt before uploading it.

      Tips

      • robots.txt controls crawling, not indexing. Use noindex to keep a page out of search results.
      • Paths are case-sensitive: /Blog/ and /blog/ are different.

      Limitations

      • Crawl-delay and non-standard directives are reported but not simulated.
      • Pasted files are tested in your browser. Loading a live file asks our server to fetch https://the-site/robots.txt.

      Frequently asked questions

      What happens if there is no robots.txt?

      A missing file (404) means everything may be crawled. A server error (5xx) can make Google pause crawling the site.

      Does Google support Crawl-delay?

      No. Google ignores Crawl-delay; Bing and some other crawlers respect it.

      Robots.txt Generator

      Create a clean robots.txt file with sitemap references, safe WordPress defaults and optional rules for AI crawlers, then…

      Sitemap URL Extractor

      Extract every URL from an XML sitemap or sitemap index, including gzipped sitemaps, and export the list to…

      Bulk Noindex Checker

      Find out which of up to 50 URLs are blocked from search results by noindex in the meta…

      HTTP Headers Checker

      View the HTTP response headers of any URL and check caching, compression and common security headers such as…

      HTTP Status Code Checker

      Check the HTTP status code a URL returns (200, 301, 404, 500…), how long it took and where…