Sitemap Finder
Enter a domain to pull every URL listed in its sitemap.xml, falling back to the Sitemap: line in robots.txt if needed.
Reads the domain's own sitemap.xml (falling back to the Sitemap: line in robots.txt), including each URL's own lastmod/priority/changefreq where the sitemap declares them — not every site sets these, so "—" means the sitemap simply didn't include that field, not that the tool failed to read it. Capped at 5,000 URLs and 20 sub-sitemaps per lookup.
Need the opposite? Use What Is My User Agent
What is Sitemap Finder?
A Sitemap Finder reads a website's own sitemap.xml file and lists every URL it declares, without you having to open the file yourself and dig through raw XML. Most sites publish a sitemap specifically so search engines (and tools like this one) can discover every page without crawling the whole site link by link. This tool fetches the domain's sitemap.xml directly. If a site doesn't have one at that path, it checks robots.txt for a Sitemap: line instead, since that's the other place sites are expected to point to it. If the sitemap is actually a sitemap index (a sitemap of sitemaps, common on larger sites), it follows those links too and combines every URL into one flat list. The result is only as complete as the site's own sitemap — a site with no sitemap, or an out-of-date one, won't return a full picture of what's actually live on the domain.
How to Use the Sitemap Finder
- 1
Enter a Domain: Type a domain like example.com (with or without https://) and press Find URLs.
- 2
Review the List: Every URL from the site's sitemap.xml (or robots.txt fallback) appears, with a live filter box to narrow it down.
- 3
Export or Copy: Use the Copy all menu to copy the list as text, CSV, or JSON, or download it as a .txt, .csv, or .json file — all based on whatever filter, sort, or category you currently have applied.
Why a Browser Can't Do This Directly
Fetching another domain's sitemap.xml or robots.txt from JavaScript running in your browser is blocked by CORS (Cross-Origin Resource Sharing) unless that specific site opts in — most don't. This tool runs the actual fetch from a server-side Cloudflare Worker instead, then hands the resulting URL list back to your browser, the same way a backend service would.
What This Tool Doesn't Do
This isn't a full site crawler — it only reads what the site's sitemap.xml (or sitemap index) and robots.txt already declare. Pages that exist but aren't listed in the sitemap won't show up here. Lookups are also capped (up to 5,000 URLs and 20 sub-sitemaps per domain) to keep each request fast and bounded, so an unusually large site's sitemap may be returned partially rather than in full.
Frequently Asked Questions
Common questions about Sitemap Finder.