Find the pages your site buries
Paste a URL. sitemap.digital crawls the site and shows its real shape: orphan pages that nothing links to, pages more than four clicks from the homepage, and the ten pages the structure treats as most important. No signup, results in seconds.
Free scan covers up to 20 pages, 50 with your emailNo signup needed
One scan, three lenses on your site's health
Every page gets checked for classic SEO fundamentals, AI crawler access, and structural integrity, all in one pass.
SEO and AI visibility
Every page gets an AI-readiness score alongside the classic SEO checks: title, meta, canonical, schema, and which AI crawlers are allowed in.
Migrations and redesigns
Compare structure before and after a move, and catch orphaned pages and broken hierarchy before they cost you traffic.
IA, UX and dev handoffs
Understand an unfamiliar site in one pass, then share the map with your team or client from a single link.
Paste a URL. Watch the map build itself.
No installs, no config. The crawl starts the moment you submit, and results stream in as each page is checked.
Live site crawl
When you paste a URL, sitemap.digital reads the sitemap and crawls up to 20 pages from it, homepage first, in the order the sitemap lists them. A site with no usable sitemap is crawled breadth-first from the homepage instead. Entering your email unlocks a 50-page scan of the same site.
Per-page analysis
For every page, the scan reads the title tag, meta description, H1 heading, canonical link, and robots meta rules, then flags anything that is missing, duplicated, or likely to hurt how the page is understood.
AI-readiness score
Alongside the classic SEO checks, each page is checked for structured data and AI-crawler access, including GPTBot, ClaudeBot, PerplexityBot, and Google-Extended.
Built for the age of AI search
Is your site readable by AI?
A page being indexed by Google is not the same as an AI system actually citing it. The AI-readiness score combines crawler access, llms.txt presence, canonical tags, structured data, and clean internal linking into one number, so you know if ChatGPT, Claude, and Perplexity can actually find and cite your pages.
See how the score is calculatedWho it is for
SEOs use sitemap.digital to get a fast, visual read on any site's structure before a deeper audit: it surfaces thin sections, broken internal links, and missing metadata in one pass instead of clicking through pages one at a time. It works just as well on a competitor or client site as on your own.
Migration and redesign teams paste the old and new URLs to compare structure before and after a move, spotting orphaned pages and broken hierarchy before launch. IA, UX, and development teams use it to understand an unfamiliar codebase or content set instantly, then share the map in a single link.
Indie developers and content teams use it as a pre-launch or pre-deploy check: paste the staging or production URL, confirm every page has the essentials it needs, and catch AI-crawler blocks before they cost visibility in AI search results.
Single-purpose checks and cornerstone guides
Free tools for when you need one answer fast, plus guides covering every check sitemap.digital runs on a scan.
Free tools
AI crawler access checker
Paste a URL and see instantly whether GPTBot, ClaudeBot, PerplexityBot, and 5 other crawlers can access your site.
Live dataThe state of AI crawlability
Live aggregate data from every site scanned here: how many have orphan pages, block AI crawlers, or ship no llms.txt.
Learn: guides on AI crawlability
Orphan pages in SEO: what they are and how to fix them
What an orphan page is, how sitemap.digital finds one from real crawl data, and the fastest way to fix it before your next scan.
GuideClick depth in SEO: what it is and why it matters
What click depth means in SEO, why pages more than four hops from the homepage get flagged, and how to flatten a site so more of it gets crawled.
GuideInternal links and the internal link graph for AI crawlers
How the internal link graph controls what GPTBot, ClaudeBot, and PerplexityBot can find on your site, and the audit checklist to fix it.
Guidellms.txt: what it is and whether it matters
The llms.txt format explained: what it is meant to do, a real worked example, and an honest, evidence based look at whether it currently changes anything.
GuideAI crawlability score: how it is calculated
How sitemap.digital scores a page for AI crawlability out of 100: content 50, structure 20, access 10, metadata 10, schema 10.
GuideWhy your site is not showing up in ChatGPT and AI answers
Your pages are indexed in Google, yet ChatGPT and other AI answers never mention your site. Here is why that gap exists and what to check.
GuideBest AI Search Visibility and Monitoring Tools (2026)
Compare 9 AI search visibility tools spanning free crawlability checks to brand mention trackers like Profound, Otterly and Peec. Honest pricing and limits.
GuideHow to Track Brand Mentions in AI Search (Manual + Tools)
A practical method for tracking whether ChatGPT, Perplexity and Google AI Mode mention your brand. Manual prompts, metrics, tools and a monthly cadence.
GuideAI Crawlers: The Complete List of Major AI Bots (2026)
Every major AI crawler explained by vendor and job: GPTBot, ClaudeBot, PerplexityBot, Bytespider, CCBot and Google-Extended. Includes a robots.txt example.
GuideClaudeBot: Anthropic's Crawler and How to Control It
ClaudeBot is Anthropic's training crawler. Learn its user agent, how it differs from Claude-SearchBot and Claude-User, and how to block it safely.
GuideGPTBot Explained: OpenAI's Crawlers and How to Control Them
GPTBot, OAI-SearchBot and ChatGPT-User are three separate OpenAI crawlers with different jobs. See the UA tokens and the robots.txt rules that control each one.
GuidePerplexityBot: Perplexity's Crawler and How to Control It
Learn what PerplexityBot is, how to verify it with IP ranges, and how to block it in robots.txt without losing Perplexity's citation traffic.
GuideBytespider: ByteDance's AI crawler and how to block it
Bytespider is ByteDance's AI crawler with no published documentation, verification or IP ranges. Learn what it does and how to block it reliably.
Frequently asked questions
Is it free?
Yes. A scan of up to 20 pages is free and does not need an account. The scan results page, the shareable link, and comparing against a later scan are all free too. An email unlocks a bigger 50-page scan, or lets you export the audit as a CSV or PDF report.
How many pages does it scan?
A free scan crawls up to 20 pages from the URL you paste, following internal links the way a normal visitor or search bot would. Entering your email unlocks a 50-page scan of the same site. Larger sites are covered breadth-first so the most important pages are checked first.
Do I need to sign up?
No. Paste a URL and the scan starts straight away, no account and no email needed. Viewing the results, sharing the link, and comparing a later scan are all free. An email is only requested if you want a bigger 50-page scan or to export the CSV or PDF report.
What is an AI-readiness score?
It is a score built from the checks that decide whether AI systems can read and cite your pages: crawler access, an llms.txt file, canonical tags, structured data, and clean internal linking. A low score means AI tools may be missing pages you actually want them to see.
Which AI crawlers do you check?
The scan looks at robots rules for the major AI crawlers, including GPTBot, ClaudeBot, PerplexityBot, and Google-Extended, alongside the standard search engine bots, so you can see at a glance who is allowed in and who is blocked.
Scan your site in seconds
No signup, no config. Paste a URL and see the full structure, SEO essentials, and AI-readiness score for up to 20 pages, 50 with your email.