Orphan pages, click depth, link graph

Find the pages your site buries

Paste a URL. sitemap.digital crawls the site and shows its real shape: orphan pages that nothing links to, pages more than four clicks from the homepage, and the ten pages the structure treats as most important. No signup, results in seconds.

Free scan covers up to 20 pages, 50 with your emailNo signup needed

SitemapLive
SitemapCRAWLING
//blog/products/about/contact/blog/category/products/reviews/careers/faq/faq/billing
sitemap.digital crawl log
20 pagesper free scan, 50 with email
No signupto start scanning
11 AI crawlerschecked per page
Secondsnot minutes
What you get

One scan, three lenses on your site's health

Every page gets checked for classic SEO fundamentals, AI crawler access, and structural integrity, all in one pass.

SEO and AI visibility

Every page gets an AI-readiness score alongside the classic SEO checks: title, meta, canonical, schema, and which AI crawlers are allowed in.

Migrations and redesigns

Compare structure before and after a move, and catch orphaned pages and broken hierarchy before they cost you traffic.

IA, UX and dev handoffs

Understand an unfamiliar site in one pass, then share the map with your team or client from a single link.

How the scan works

Paste a URL. Watch the map build itself.

No installs, no config. The crawl starts the moment you submit, and results stream in as each page is checked.

01 / CRAWL

Live site crawl

When you paste a URL, sitemap.digital reads the sitemap and crawls up to 20 pages from it, homepage first, in the order the sitemap lists them. A site with no usable sitemap is crawled breadth-first from the homepage instead. Entering your email unlocks a 50-page scan of the same site.

02 / CHECK

Per-page analysis

For every page, the scan reads the title tag, meta description, H1 heading, canonical link, and robots meta rules, then flags anything that is missing, duplicated, or likely to hurt how the page is understood.

03 / SCORE

AI-readiness score

Alongside the classic SEO checks, each page is checked for structured data and AI-crawler access, including GPTBot, ClaudeBot, PerplexityBot, and Google-Extended.

The AI-readiness score

Built for the age of AI search

example.comIllustrative example
82
AI readiness score
Content
30
Access
25
Structure
20
Metadata
15
Schema
10

Is your site readable by AI?

A page being indexed by Google is not the same as an AI system actually citing it. The AI-readiness score combines crawler access, llms.txt presence, canonical tags, structured data, and clean internal linking into one number, so you know if ChatGPT, Claude, and Perplexity can actually find and cite your pages.

See how the score is calculated

Who it is for

SEOs use sitemap.digital to get a fast, visual read on any site's structure before a deeper audit: it surfaces thin sections, broken internal links, and missing metadata in one pass instead of clicking through pages one at a time. It works just as well on a competitor or client site as on your own.

Migration and redesign teams paste the old and new URLs to compare structure before and after a move, spotting orphaned pages and broken hierarchy before launch. IA, UX, and development teams use it to understand an unfamiliar codebase or content set instantly, then share the map in a single link.

Indie developers and content teams use it as a pre-launch or pre-deploy check: paste the staging or production URL, confirm every page has the essentials it needs, and catch AI-crawler blocks before they cost visibility in AI search results.

Free tools and guides

Single-purpose checks and cornerstone guides

Free tools for when you need one answer fast, plus guides covering every check sitemap.digital runs on a scan.

Free tools

Learn: guides on AI crawlability

Guide

Orphan pages in SEO: what they are and how to fix them

What an orphan page is, how sitemap.digital finds one from real crawl data, and the fastest way to fix it before your next scan.

Guide

Click depth in SEO: what it is and why it matters

What click depth means in SEO, why pages more than four hops from the homepage get flagged, and how to flatten a site so more of it gets crawled.

Guide

Internal links and the internal link graph for AI crawlers

How the internal link graph controls what GPTBot, ClaudeBot, and PerplexityBot can find on your site, and the audit checklist to fix it.

Guide

llms.txt: what it is and whether it matters

The llms.txt format explained: what it is meant to do, a real worked example, and an honest, evidence based look at whether it currently changes anything.

Guide

AI crawlability score: how it is calculated

How sitemap.digital scores a page for AI crawlability out of 100: content 50, structure 20, access 10, metadata 10, schema 10.

Guide

Why your site is not showing up in ChatGPT and AI answers

Your pages are indexed in Google, yet ChatGPT and other AI answers never mention your site. Here is why that gap exists and what to check.

Guide

Best AI Search Visibility and Monitoring Tools (2026)

Compare 9 AI search visibility tools spanning free crawlability checks to brand mention trackers like Profound, Otterly and Peec. Honest pricing and limits.

Guide

How to Track Brand Mentions in AI Search (Manual + Tools)

A practical method for tracking whether ChatGPT, Perplexity and Google AI Mode mention your brand. Manual prompts, metrics, tools and a monthly cadence.

Guide

AI Crawlers: The Complete List of Major AI Bots (2026)

Every major AI crawler explained by vendor and job: GPTBot, ClaudeBot, PerplexityBot, Bytespider, CCBot and Google-Extended. Includes a robots.txt example.

Guide

ClaudeBot: Anthropic's Crawler and How to Control It

ClaudeBot is Anthropic's training crawler. Learn its user agent, how it differs from Claude-SearchBot and Claude-User, and how to block it safely.

Guide

GPTBot Explained: OpenAI's Crawlers and How to Control Them

GPTBot, OAI-SearchBot and ChatGPT-User are three separate OpenAI crawlers with different jobs. See the UA tokens and the robots.txt rules that control each one.

Guide

PerplexityBot: Perplexity's Crawler and How to Control It

Learn what PerplexityBot is, how to verify it with IP ranges, and how to block it in robots.txt without losing Perplexity's citation traffic.

Guide

Bytespider: ByteDance's AI crawler and how to block it

Bytespider is ByteDance's AI crawler with no published documentation, verification or IP ranges. Learn what it does and how to block it reliably.

FAQ

Frequently asked questions

Is it free?

Yes. A scan of up to 20 pages is free and does not need an account. The scan results page, the shareable link, and comparing against a later scan are all free too. An email unlocks a bigger 50-page scan, or lets you export the audit as a CSV or PDF report.

How many pages does it scan?

A free scan crawls up to 20 pages from the URL you paste, following internal links the way a normal visitor or search bot would. Entering your email unlocks a 50-page scan of the same site. Larger sites are covered breadth-first so the most important pages are checked first.

Do I need to sign up?

No. Paste a URL and the scan starts straight away, no account and no email needed. Viewing the results, sharing the link, and comparing a later scan are all free. An email is only requested if you want a bigger 50-page scan or to export the CSV or PDF report.

What is an AI-readiness score?

It is a score built from the checks that decide whether AI systems can read and cite your pages: crawler access, an llms.txt file, canonical tags, structured data, and clean internal linking. A low score means AI tools may be missing pages you actually want them to see.

Which AI crawlers do you check?

The scan looks at robots rules for the major AI crawlers, including GPTBot, ClaudeBot, PerplexityBot, and Google-Extended, alongside the standard search engine bots, so you can see at a glance who is allowed in and who is blocked.

Scan your site in seconds

No signup, no config. Paste a URL and see the full structure, SEO essentials, and AI-readiness score for up to 20 pages, 50 with your email.