Two files.
One contradiction.

Your sitemap invites the crawler to pages your robots.txt forbids. Neither file is wrong on its own. Together, they contradict each other, and that contradiction is invisible until someone reads both at once.

How it works

  • 01
    Fetch and parse robots.txt
    Robotsdiff reads your robots.txt and extracts every User-agent group, Disallow rule, Allow directive, and Sitemap declaration.
  • 02
    Discover and fetch sitemaps
    Every sitemap URL declared in robots.txt is fetched. If none is declared, common locations are checked. Sitemap indexes are expanded one level deep.
  • 03
    Compare and report contradictions
    Every URL from every sitemap is checked against every Disallow rule. URLs that appear in both a sitemap and a Disallow rule are flagged — those are pages you asked Google to index and told it not to read.

What you get

  • --
    Contradiction report
    Every URL that appears in a sitemap and matches a Disallow rule, listed with the exact directive that blocks it and the sitemap that includes it.
  • --
    Uncrawled path report
    Paths that are crawlable per robots.txt but absent from any sitemap — informational, not a finding, because not every page belongs in a sitemap.
  • --
    Full file inspection
    View the parsed robots.txt groups and rules alongside the sitemap content, so you can see exactly what each file declares.

What it does not do

It does not judge whether a page should be indexed. A contradiction means two files disagree — it does not mean the page is broken or should be removed from the sitemap. Some Disallow rules are intentional access controls, and putting a restricted page in the sitemap may serve internal publishing workflows. The tool reports the disagreement and stops there.

It does not fetch the pages themselves or estimate ranking impact. It compares two text files and names where they differ.

It does not detect every possible contradiction. robots.txt wildcard patterns are approximated with prefix and simple glob matching. Extremely malformed sitemap XML may produce incomplete results. Sitemap indexes are followed one level deep to avoid infinite loops.

Cloudflare-hosted domains: A Worker cannot reach a host that is itself behind Cloudflare. If your domain uses Cloudflare, this tool may report it as unreachable even though the site is up. This is a platform limitation, not a problem with your server.

Check your domain

Paste any public domain. No sign-up, no credentials. Results in seconds.

Check a domain