Sitemap URL ExtractorSEO Tools
Reads sitemap files, follows sitemap indexes, and exports the URL list with lastmod values when available.
Find sitemap of website files from one domain. The tool reads robots.txt, tries the usual sitemap addresses, and lists each sitemap it finds with its type, URL count, and newest lastmod.
Reads the Sitemap lines in robots.txt, then tries the usual sitemap addresses.
Tick the box to show you are not a robot, then press Read.
example.com, or any page address on the site. Only the scheme and host are used.The server first downloads /robots.txt and keeps every Sitemap: line. Then it checks five addresses that sites often use: /sitemap.xml, /sitemap_index.xml, /sitemap-index.xml, /wp-sitemap.xml, and /sitemap.xml.gz. An address that answers with a 2xx or 3xx status counts as found.
Each sitemap found is downloaded and parsed. A file with a <sitemapindex> element is an index, and its count is the number of child sitemaps. Any other file is a URL set, and its count is the number of <url> entries. The newest lastmod value is shown for URL sets. Gzip files are unpacked first.
Every request goes to public addresses only, and the domain is not stored.
/wp-sitemap.xml as an index, with child sitemaps for posts, pages, taxonomies, and users.https://example.com/sitemap.xml in robots.txt and also answers on that path shows it once, because duplicate addresses are merged.Look for a Sitemap: line in /robots.txt, then try /sitemap.xml. This tool does both, plus four other common paths, in one check.
A URL set lists page addresses. A sitemap index lists other sitemap files, which large sites use because one file may hold at most 50,000 URLs.
It has none, or it keeps the file at an address that robots.txt does not mention. The site owner can add a Sitemap: line to robots.txt so that every crawler finds it.
No. The domain is used for this check and is not stored.
Often used together with Website Sitemap Extractor.
Reads sitemap files, follows sitemap indexes, and exports the URL list with lastmod values when available.
Build sitemap.xml files from pasted URLs or collect links from one page and export valid sitemap XML.
Creates robots.txt rules for search and AI crawlers and tests URLs against them.