Sitemap Checker

Find your sitemap from robots.txt or /sitemap.xml, validate the XML, count URLs, handle sitemap index files, and spot common issues search engines trip on.

How it works

  1. Enter your website URL, or paste a sitemap URL directly (anything ending in .xml is treated as a sitemap).

  2. For a site URL, the tool discovers your sitemap the way crawlers do: from the Sitemap line in robots.txt first, then at the default /sitemap.xml location.

  3. It validates the XML structure, counts the listed URLs, and follows sitemap index files into their child sitemaps.

  4. It then spot-checks a sample of listed URLs — do they respond, and do they live on the same domain as the sitemap?

  5. Finally it validates <lastmod> date formats and reports everything as pass/warning/fail checks with fixes.

Frequently asked questions

Does an XML sitemap guarantee that pages will be indexed?

No. A sitemap helps search engines and many AI crawlers discover URLs, but indexing still depends on crawlability, quality, and whether the page is allowed to be indexed. Blocking via robots.txt, a noindex tag, soft-404s, or thin duplicate content can keep listed URLs out of the index. Submit the sitemap in Google Search Console (and Bing Webmaster Tools) so you can track discovery and see why specific URLs are excluded.

What is a sitemap index file?

A sitemap index is a single XML file that lists other sitemaps instead of individual page URLs. Large sites use it when one sitemap would exceed URL or file-size limits, splitting content across child sitemaps (for example by section or language). Crawlers fetch the index first, then follow each child sitemap to discover the full URL set.

How important are lastmod dates in a sitemap?

lastmod is optional, but accurate dates help crawlers prioritize recrawling pages that actually changed. If dates are missing, stale, or updated on every deploy whether content changed or not, engines tend to ignore them. Use W3C datetime format and set lastmod only when the meaningful content of that URL changed.

Why do broken URLs in a sitemap hurt SEO?

A sitemap is a strong signal that those URLs are worth crawling. When entries return 404s, 5xx errors, or redirects to irrelevant destinations, crawlers waste budget on dead ends and may trust the sitemap less over time. Keep the file limited to live, canonical, indexable URLs and remove or fix broken listings promptly.

Do AI crawlers use XML sitemaps the same way search engines do?

Many AI and answer-engine crawlers discover URLs via the same paths search bots use: robots.txt Sitemap lines, /sitemap.xml, and linked sitemap indexes. An XML sitemap is the crawler-facing format; an HTML sitemap for humans is useful for visitors but does not replace a valid XML sitemap for automated discovery. Keeping XML sitemaps current still improves chances that both traditional search and AI crawlers find your important pages.

Get help fixing this

Fortitude builds fast, search-friendly websites. Leave your email and we'll help you fix what's holding your site back.

More free tools

SERP Snippet Preview

Preview how your title and meta description will look in Google search results.

Meta Description Checker

Check any page’s meta description for length, clarity, and click-worthiness.

H1 Checker

Find missing, duplicate, or weak H1 headings on any page.

Canonical Checker

Verify canonical tags are present, absolute, and pointing to the right URL.

Robots.txt Checker

Check your robots.txt for syntax issues and crawler access, including AI bots.

Robots.txt Generator

Generate a clean robots.txt with sitemap and AI-crawler rules in seconds.

llms.txt Validator

Validate your llms.txt structure so AI assistants can understand your site.

llms.txt Generator

Generate a draft llms.txt from your site’s pages and description.

AI SEO Audit

A full AI-readiness SEO report: crawlability, metadata, structure, and fixes.

GEO Audit

Check how ready a page is for AI search and answer engines — not just Google.

Schema Markup Checker

Find and validate the JSON-LD structured data on any page.

AI Readability Checker

Check how easily AI assistants can parse, summarize, and quote your page.

FAQ Generator

Generate FAQ questions and answers from your page content, ready for FAQPage schema.

Citation Readiness Checker

Check how clearly your page states facts an AI assistant could reference.

Entity Clarity Checker

Check how clearly your page identifies who or what it is about.

AI Overview Readiness

Check how ready a page is to answer a specific query in an AI-generated answer block.

PageSpeed & Core Web Vitals Checker

Check Core Web Vitals — LCP, INP, and CLS — for any page via Google PageSpeed Insights.

Get A Quote