Robots.txt & Sitemap Checker

Make sure search engines can crawl what matters: we read robots.txt, test a URL against Googlebot’s rules and check the XML sitemap it points to.

  • Googlebot rules
  • Fast results
  • 100% free
  • No signup

Add a page path to test whether Googlebot may crawl it · free

How it works

How to use the robots.txt & sitemap checker

  1. 1

    Enter a URL

    Use your home page, or a specific page to test whether it’s blocked.

  2. 2

    We read the rules

    robots.txt is parsed with Google’s matching rules and the sitemap is loaded.

  3. 3

    Unblock & submit

    Fix accidental blocks and submit the sitemap in Google Search Console.

What robots.txt does (and doesn’t do)

robots.txt is a plain-text file at the root of your site that tells crawlers which paths they may crawl. It does not remove pages from Google — a blocked URL can still be indexed if other sites link to it. To keep a page out of search results, use a noindex meta tag and leave it crawlable.

How Google reads the rules

Google picks the most specific User-agent group that matches its crawler, then applies the longest matching Allow or Disallow rule; when they tie, Allow wins. * matches any characters and $ anchors the end of the URL. This checker uses the same logic, so you can test any path.

Why your sitemap belongs in robots.txt

An XML sitemap lists the URLs you want indexed, with optional last-modified dates. Adding a Sitemap: line to robots.txt lets Google, Bing and every other search engine discover it automatically — even if you never submit it manually.

FAQ

Robots.txt & Sitemap Checker FAQ

Does every site need a robots.txt file?

No. Without one, search engines crawl everything. It’s still recommended so you can list your sitemap and keep crawlers out of areas like admin or cart pages.

Will blocking a page in robots.txt remove it from Google?

No. It only stops crawling. Use a noindex tag (and keep the page crawlable) or the removal tool in Search Console to drop a page from results.

How big can a sitemap be?

Up to 50,000 URLs or 50 MB uncompressed per file. Larger sites use a sitemap index that points to several sitemaps.

Is Crawl-delay supported?

Bing and some other crawlers respect it; Googlebot ignores it and adjusts its crawl rate automatically.