Free online tool

Sitemap Extractor

Enter a domain and get every URL the site publishes in its sitemap — grouped by section, searchable, and ready to export or turn into Markdown.

Try
What it reads
  1. robots.txt
  2. sitemap.xml
  3. sitemap_index.xml
  4. .xml.gz
  5. On-page links fallback
How it works

The sitemap the site publishes, not a guess

  1. Enter a domain

    A bare domain like example.com is enough; any URL on the site works the same way. Only URLs on that host are kept.

  2. Sitemaps are discovered

    Sitemap: lines in robots.txt win, then the conventional sitemap paths. Sitemap indexes and gzip files are expanded automatically.

  3. Browse, filter and export

    URLs are grouped by top-level section. Filter by keyword, export TXT or CSV, or send any page to the Markdown converter.

Use cases

What people use a full URL list for

SEO audits

Check what you actually submit to search engines: stray staging pages, missing sections, or URLs that should not be indexed.

Site migrations

Export the old site’s URL inventory as CSV and use it as the source list for redirect mapping before you switch.

Crawling & RAG

Get the exact page list to crawl instead of following links blindly, then convert each page to Markdown for your index.

Content research

See how a competitor organizes docs, blog and landing pages, and how many pages each section holds.

Frequently asked questions

Where does the URL list come from?

From the XML sitemaps the site publishes. Sitemap: directives in robots.txt are read first, then /sitemap.xml, /sitemap_index.xml, /sitemap-index.xml and /sitemap.xml.gz. Sitemap indexes are expanded one level.

What if the site has no sitemap?

The extractor falls back to the links found on the page you entered, filtered to the same host. That list is usually much shorter than a real sitemap.

Is there a limit on the number of URLs?

One extraction returns up to 5,000 URLs, in the order the sitemaps list them. Subdomains such as docs.example.com have their own sitemaps and can be extracted separately.

Is it free?

Yes. Guests get 10 extractions a day. With a free account each extraction costs 1 credit from your own API key, and new accounts start with 100 credits.

Can I export the results?

Yes. Copy every URL to the clipboard, or download a .txt list or a .csv file with the URL, path and section of each page.

Can I use this from code?

The page runs on the Search1API /sitemap endpoint. The same request works from the REST API, the SDKs and the MCP server.

Map sites from your own code

The same sitemap extraction is one POST request away. Pair it with /crawl to turn every page into Markdown.