The gap
Sitemaps go stale quietly: pages get redirected, removed or set to noindex, canonicals point elsewhere and hreflang pairs lose their return links, and search consoles report it weeks later, one sample at a time.
Developer tools
A command-line sitemap auditor that requests every URL in your XML sitemaps and lists the ones search engines will trip over: errors, redirects, noindex pages, wrong canonicals and broken hreflang pairs.
Coming soon
Not published on its store yet, so there is no buy or download link. PN Scripts does not take payment for it here.
The gap
Sitemaps go stale quietly: pages get redirected, removed or set to noindex, canonicals point elsewhere and hreflang pairs lose their return links, and search consoles report it weeks later, one sample at a time.
What it is
Mapwarden checks every URL in the sitemaps in one run, explains each problem with a code and a message, and returns an exit code a deployment pipeline can act on.
How you use it
Find every broken, redirected or contradictory sitemap entry before a search engine does, with a report you can diff, filter and open in a spreadsheet.
For developers and technical SEO people who maintain multilingual or fast-changing sites, Mapwarden is an offline, scriptable sitemap audit that respects robots.txt and fits into CI.
Use it when
Built for
Point Mapwarden at a site root and it finds the sitemaps through robots.txt, follows indexes and gzip files, removes duplicates and checks each URL. A summary groups the findings by code, so the biggest problem is the first line you read.
HTTP errors and failed requests, redirects, chains of more than one hop and loops, noindex in meta robots or the X-Robots-Tag header, canonical links that point elsewhere or conflict, missing, short or long titles, URLs blocked by robots.txt, duplicates and URLs on other hosts.
Every alternate is resolved and compared against the other audited pages. Missing return links, invalid codes such as en-UK, duplicate languages, alternates that redirect or fail and sets without a self-reference are reported on the page that has the problem.
robots.txt rules and Crawl-delay are obeyed, with 4 parallel requests and 2 requests per second per host by default. JSON and CSV reports, the -only filter and -fail-on exit codes make it easy to run in CI or on a schedule.
1 of 5
| Platform or runtime | Supported versions | Tested up to |
|---|---|---|
| Linux amd64 (binary run) | kernel 3.2 or newer | Ubuntu 24.04, kernel 6.8 |
| Go (to build from source) | 1.22 or newer | 1.26.0 |
It is finished and tested. It is not on sale yet; this page will say where to get it when it is.
Coming soon
Not published on its store yet, so there is no buy or download link. PN Scripts does not take payment for it here.