Sitemap Frog: Sitemap Generator & Auditor
Overview
Crawl and audit any site you operate. Review discovered URLs, fix SEO issues, and export a perfect XML sitemap.
🐸 Sitemap Frog crawls a website you operate, audits every page it finds, and turns the result into a sitemap you control. Unlike a server-side crawler, it runs in your browser, so it can render JavaScript pages — React, Vue, Angular and other single-page apps that serve an empty shell to ordinary crawlers are followed correctly. 🔎 WHAT IT FINDS Links, of course, but also: • Link relations (canonical, next/prev, hreflang) • Iframes • Image srcsets and lazy-loaded images • CSS url() references • JSON-LD • Data attributes • Framework data payloads • RSS and Atom feeds • HTTP Link headers • Redirect targets • Any sitemap the site already publishes Underneath: • Titles • Meta descriptions • H1 and H2 • Structured data • Response headers • Anchor text • Response times • The full internal link graph 🩺 WHAT IT AUDITS Every crawl ends with a 153-finding technical SEO and accessibility audit across 23 areas — response codes, URL shape, titles, meta descriptions, meta keywords, H1, H2, content, images, canonicals, pagination, directives, hreflang, security, links, JavaScript dependence, structured data, sitemaps, accessibility, mobile usability, markup, AMP and PageSpeed. Highlights: • Duplicate titles, descriptions, H1, H2 and keywords, grouped across the crawl • Exact duplicates by hash, near duplicates by minhash under locality-sensitive hashing • Link score — damped PageRank over your internal links, log-scaled 0 to 100 • Orphan pages, redirect and canonical chains and loops, hreflang return-link mismatches • Pagination sequence errors and oversized images • Around twenty WCAG 2.1 A and AA rules, plus colour contrast, tap target size and font size measured on rendered pages • Structured data validated against Google's rich-result requirements for 27 schema types • AMP markup checks on pages that declare it Every finding explains itself — the consequence if you leave it, why it matters, how to fix it, and the actual evidence on the page (the missing alt texts, the insecure URLs, the failing contrast ratios, the other pages sharing a title). Click a finding to filter the URL table to exactly the affected pages. Re-run audit re-scores a stored crawl without re-fetching. 🔌 OPTIONAL EXTERNAL DATA • Google PageSpeed Insights — supply your own API key to pull Lighthouse scores and Core Web Vitals straight into the audit. Off by default. Sends selected URLs and your key directly to Google. • Moz backlink metrics — supply your own API token to attach Domain Authority, Page Authority, Spam Score and referring domains to your URLs. Off by default. Sends selected URLs and your token directly to Moz. ⚙️ WHAT YOU CONTROL Review every discovered URL in a searchable table. Include or exclude individually, by selection, or by filter. Crawl a list of URLs instead of spidering from a seed. Compare the current crawl against a JSON backup of an earlier one to see what appeared, disappeared or changed. Inspect a single page with raw HTML side-by-side against the rendered DOM, with an optional screenshot. Visualise the crawl as: • A directory tree • A crawl tree • A force-directed link graph • Anchor-text and content word clouds Click any node to filter the table to that URL. Test your robots.txt against any URL list before pushing it live. Set lastmod, changefreq and priority globally, per depth, or per URL. 📤 EXPORT AS • 📄 Sitemap XML — automatically splits into a sitemap index past 50,000 URLs • 🖼️ Image sitemap • 🎬 Video sitemap • 📝 Plain text • 📊 CSV • 🌐 HTML sitemap page • 🧩 JSON backup • 📑 38 focused reports (redirects, chains, orphans, duplicates, hreflang, structured data, PageSpeed, anchor text, n-grams, custom search, custom extraction…) • 📈 One .xlsx workbook containing everything 🛡️ IT PLAYS BY THE RULES robots.txt, rel="nofollow", meta robots and X-Robots-Tag are all respected by default. Full Google-spec robots.txt matching — wildcards, $ anchors, longest match, per-subdomain files, Crawl-delay. 🔒 REQUESTS & COOKIES Requests are made without your cookies by default, so Sitemap Frog only sees what an anonymous visitor sees. Optionally enable cookies to crawl as your signed-in session. 🌍 AVAILABLE IN 54 LANGUAGES The full interface, not just the store listing. 🔐 PRIVACY Nothing is transmitted anywhere by default. ✓ No account ✓ No server ✓ No analytics Crawl results stay in your browser and you can delete them at any time. The extension fetches a site only when you start a crawl of it — without sending your cookies. PageSpeed Insights and Moz backlink metrics are off by default and send selected URLs plus your own credential directly to Google or Moz only when you enable them. For a normal `<textarea>`, this can be pasted directly as-is. The emojis will remain plain Unicode text rather than requiring HTML.
0 out of 5No ratings
Details
- Version0.3.0
- UpdatedSeptember 18, 2026
- Offered byHintaro
- Size669KiB
- Languages54 languages
- DeveloperBSK
B-III, 1203, Bajwa Nagar, Main Road Ludhiana, Punjab 141001 INEmail
anmolhost@gmail.com - Non-traderThis developer has not identified itself as a trader. For consumers in the European Union, please note that consumer rights do not apply to contracts between you and this developer.
Privacy
This developer declares that your data is
- Not being sold to third parties, outside of the approved use cases
- Not being used or transferred for purposes that are unrelated to the item's core functionality
- Not being used or transferred to determine creditworthiness or for lending purposes