Free Sitemap URL Counter & XML Analyzer

Count URLs in an XML sitemap or nested sitemap index. Enter a public URL, paste raw XML, or load an XML/XML.GZ file. Local files stay in your browser. See which top-level paths contain the most pages, browse a visual sitemap tree, and check folder depth. Measure duplicate URLs and metadata coverage, chart the declared lastmod age, then inspect lastmod, changefreq, and priority in a filterable table. Include or exclude URL patterns, then copy or download the filtered list as CSV, TXT, or JSON.

Count URLs in a Sitemap Index

This sitemap URL counter parses regular XML sitemaps and follows the child files in a sitemap index. The result separates the total URL count from the number of sitemap files processed, so an index with 20 child files is not mistaken for a sitemap containing only 20 pages.

Gzipped .xml.gz files are decompressed automatically, including compressed child files linked from a sitemap index. You can use the same URL field for plain XML and gzip-compressed sitemaps.

The extractor also returns the URLs themselves. Each entry can include optional metadata defined by the sitemap protocol: last modification date, change frequency, and priority.

Extracting sitemap data is useful for SEO audits, content inventories, site migration planning, and competitor analysis. Our tool handles both regular sitemaps and sitemap index files, following references automatically to give you a complete URL list. If you would rather use a script or command line, see our guide to extracting sitemap URLs.

How the Sitemap Extractor Works

1

Fetch & parse

We fetch the sitemap URL and parse the XML to extract all <url> entries with their metadata.

2

Follow index files

If the URL points to a sitemap index, we automatically follow all referenced sitemaps and combine the results.

3

Filter & download

Keep URLs matching one pattern, remove URLs matching another, and download the filtered list as CSV, TXT, or JSON.

Frequently Asked Questions

What is a sitemap extractor?

A sitemap extractor reads an XML sitemap file and pulls out all the URLs it contains, along with metadata like last modified dates, change frequency, and priority values. This is useful for SEO audits, content inventories, and migration planning.

How many URLs can I extract?

The free tool extracts up to 5,000 URLs per sitemap. If the sitemap contains more, results are truncated. A free API key includes 1,000 URLs per extraction. Paid plans support up to 50,000.

Can I download the extracted URLs?

Yes. You can download the extracted URLs as a CSV file (with all metadata columns), a plain text file (URLs only), or a JSON file. All download formats are available directly from the results.

Does it follow sitemap index files?

Yes. If the URL you provide points to a sitemap index file (<sitemapindex>), the tool automatically follows the referenced sitemaps and extracts URLs from all of them.

Does it support gzipped sitemaps (.xml.gz)?

Yes. Enter the .xml.gz URL as you would a regular sitemap URL. The extractor decompresses the file before parsing it and also handles gzip compression sent through the HTTP Content-Encoding header.

Can I include or exclude URLs before downloading?

Yes. Use the include field to keep URLs containing a path or word, and the exclude field to remove unwanted matches. The table, copied URLs, and CSV, TXT, or JSON downloads all use the filtered list.

Need to extract sitemaps at scale?

Get an API key to automate extraction. Paid plans support up to 50,000 URLs per sitemap. 20 credits/month included.