Free tool
XML Sitemap Generator
Enter any website address. We crawl its internal links live, skip pages that should not be indexed, and hand you a valid sitemap.xml to download or copy.
Why generate a sitemap from a crawl?
A sitemap built from a live crawl reflects what a search engine can actually reach by following your links, which is precisely the point. Pages that exist in the CMS but are linked from nowhere will not appear, and that is a feature, not a bug: if our crawler cannot find a page, Googlebot will struggle too. Fix the internal link first, then regenerate.
What the generator checks on every page
For each discovered URL the tool records the HTTP status, honours noindex meta tags and X-Robots-Tag headers, reads the Last-Modified header where the server provides one (it becomes the <lastmod> value), and normalises the URL by stripping fragments and query strings. The output is standards-compliant XML that validates against the sitemaps.org schema, ready for Search Console.
When you should not use a generated file
If your site runs on a CMS that can serve a dynamic sitemap, use that instead: it updates itself when you publish. A hand-uploaded static file is right for hand-coded HTML sites, legacy stacks with no plugin ecosystem, and quick audits where you want to count reachable pages. It is also a fast way to see how much of a site survives after a migration; compare the URL count before and after.
Sitemap myths worth ignoring
Submitting a sitemap does not guarantee indexing; Google treats it as a hint, and low-quality pages stay out regardless. Priority and changefreq fields are ignored by Google, which is why this generator does not emit them. And a sitemap does not need to list every URL variant, only the canonical ones you actually want in search.
Frequently asked questions
Which pages end up in the sitemap?
Pages of the same site that answer HTTP 200 with an HTML content type and carry no noindex directive. Files (images, PDFs, scripts), URLs with query strings, and service paths like /wp-admin/ or /cart/ are excluded automatically, because they do not belong in a sitemap.
Why is the tool limited to 80 pages?
A live crawl costs real requests to your server, and 80 pages covers the vast majority of business sites, landing pages and portfolios. If your site is bigger, your CMS should generate the sitemap natively (WordPress does since 5.5, and SEO plugins do it better still); a static file would go stale anyway.
Does a sitemap improve my rankings?
Not directly. A sitemap helps crawlers discover URLs faster and report indexing status per file in Search Console. It is plumbing, not promotion: important for new, large or poorly linked sites, nearly irrelevant for a 10-page site that is already well interlinked.
Where do I put the file once I download it?
Upload it to the site root so it is reachable at yourdomain.com/sitemap.xml, reference it in robots.txt with a Sitemap: line, and submit it once in Google Search Console and Bing Webmaster Tools. After that, keep it updated when pages change.
Ready to get an exact copy of any website?
Send us a link and get a free, no-obligation estimate within one business day.
Get a free quote