Find any website's sitemap

Enter a domain. We read its robots.txt the way a crawler does, and if it declares no sitemap we probe the conventional paths. You get every sitemap found, where it was found, and how many URLs it leads to.

Free, no signup, no limits worth mentioning.

How it finds them

It asks robots.txt first. The Sitemap: directive is site-wide and case-insensitive, valid anywhere in the file, so every line is scanned rather than only the lines inside a matching User-agent group — a rule plenty of parsers get wrong.

Only if robots.txt names nothing does it guess, probing the conventional paths a crawler would try. That distinction is reported back to you: “we found it where you declared it” and “we had to guess” are different answers about how well your site is configured.

What you get back

Every sitemap discovered, up to ten, with the URL it was found at, whether it came from robots.txt or a probe, and how many URLs it contains. Sitemap indexes are expanded so the count reflects real pages rather than the number of child files.

Why it matters for a chatbot

A sitemap is the cheapest complete list of your own pages that exists. Without one, anything indexing your site — a search engine, or a chatbot trained on your content — has to discover pages by following links, and anything not linked from somewhere obvious is simply never found.

Common questions

What is a sitemap, in one sentence?

An XML file listing the pages on a site that its owner wants crawlers to know about, usually at /sitemap.xml and usually declared in robots.txt.

Why does this find sitemaps that Google does not show me?

Search Console only reports sitemaps that have been submitted to it. This reads what your site actually publishes right now, which is often more, and occasionally an old one nobody remembered was still there.

It found nothing. Does that mean I have no sitemap?

It means there is none declared in robots.txt and none at any of the conventional paths. A sitemap can exist at an unusual URL and still work, provided something points crawlers at it, but if nothing does then in practice it is not being used.

Do you store the sites people check?

No. The scan runs when you press the button and the result is returned to your browser. Nothing about it is kept afterwards, and no signup is involved.

Now let it answer your visitors' questions

TeklTalk reads the same pages you just checked and answers questions about them, with a link back to the page each answer came from.

Free to start. No credit card required.