Foundational10 min read

Crawl depth explained

Crawl depth is how many clicks it takes to reach a page from your homepage — the shortest path through internal links. It's one of the most predictive structural metrics there is, because depth correlates tightly with two things that decide rankings: how often a page is crawled and how much internal authority it receives. This is the concept explained: what depth is, why it matters so much, how it relates to authority, and what a healthy distribution looks like. For the step-by-step fixes, see the companion how-to.

Run the Crawl Depth Checker on your site — free, no account.

Check your crawl depth free

What crawl depth measures

Crawl depth is the click-distance from the homepage (depth 0) to a page, following the shortest chain of internal links. The homepage's direct links are depth 1, what they link to is depth 2, and so on. Crucially it's about clicks, not URL folders: a page at /a/b/c/d/e is depth 1 if the homepage links straight to it, and a page at /article can be depth 6 if it's only reachable through five pagination hops.

Depth is click-distance, not URL nesting
[Home] depth 0
   |
   ├─► /products        depth 1
   │      └─► /products/shoes   depth 2
   │             └─► /products/shoes/running  depth 3
   │
   └─► /blog/post-from-2019      depth 1
          (homepage links to it directly, so
           it's shallow despite the URL)
Depth follows the link graph, not the address bar. Anything the homepage or a hub links to directly is shallow, regardless of how deep its URL path looks.

How depth is actually computed

Depth is a shortest-path calculation over your internal link graph — a breadth-first walk out from the homepage, where each page's depth is the fewest links needed to reach it. Because it is the shortest path, one good link is enough to make a page shallow no matter how buried it is elsewhere.

That also means depth is only as accurate as the link graph you measured, and several things quietly change the answer:

  • JavaScript navigation — if the menu is not in the server-rendered HTML, a text-only crawl never sees those edges and reports the whole site as far deeper than it is.
  • Nofollowed internal links — whether they count depends on the crawler's configuration, so two tools can legitimately disagree.
  • The starting URL — depth from the homepage is the convention, but a site where most entry is from search has a different practical shape.
  • Crawl budget — if the crawl stopped early, pages beyond the cut appear infinitely deep rather than deep.
  • Redirects — a link through a 301 is still an edge, but chained redirects add hops that some tools count and others collapse.

Before concluding your site is too deep, confirm the crawl rendered JavaScript. An implausible depth distribution is far more often a rendering-mode artefact than a real structural problem.

Why depth decides crawling and ranking

Two independent mechanisms make depth matter, and they compound:

  • Crawl frequency falls with depth. Search engines crawl shallow, well-linked pages more often, because they look more important and are reached sooner in the crawl. A deep page is recrawled rarely, so its updates are indexed late and it can drift stale.
  • Authority falls with depth. Because internal authority decays with every hop, a deep page inherits only a heavily-reduced share — the same idea as authority travel distance. Depth and authority are two readings of one underlying problem.

The practical line: Keep pages that matter within about three clicks of a strong entry point. Past that, both crawl frequency and authority drop off sharply, and good content starts underperforming for reasons that have nothing to do with the content.

Pagination is the main way depth accumulates

Most deep tails are not built deliberately — they are produced by pagination, one publication at a time.

A blog index links the ten most recent posts. Post eleven moves to page 2, which is itself depth 2, making that post depth 3. By post ninety it is on page 9 — reachable only after eight pagination clicks, so depth 10 despite nothing about it changing. The page did not get worse; the conveyor belt moved.

How a post sinks without anyone touching it
  week 1    /blog  ->  post          depth 2\n  month 2   /blog  ->  page/2  ->  post   depth 3\n  month 6   /blog  ->  page/5  ->  post   depth 6\n  year 1    /blog  ->  page/12 ->  post   depth 13\n
Same post, same content, same links pointing at it. Only the archive moved.
  • Ecommerce — the same effect on category pagination, made worse when out-of-stock products drop out of the default listing entirely.
  • Deep filters — a product reachable only by selecting three facets in sequence is effectively very deep, because a crawler will not guess the combination.
  • Single-path sections — a whole subsection hanging off one navigation link inherits that link's depth, and loses everything if the link is removed.
  • Documentation — versioned or deeply nested docs where the useful page sits five levels below the entry point.

The structural answer to all of them is the same: stop relying on sequential paths for discovery, and build hubs that link directly to everything in a topic. That is what content clusters do — they collapse a chain into two hops.

What a healthy depth distribution looks like

Depth is best read as a distribution — how many pages sit at each click-depth — not a single number. A healthy site has most of its important pages shallow and only low-value pages in the deep tail.

Healthy vs unhealthy depth profiles
HEALTHY                       UNHEALTHY
depth  pages                  depth  pages
  0    █                        0    █
  1    █████                    1    ██
  2    ████████  <- bulk of     2    ███
  3    ████         valuable    3    █████
  4+   ██  (minor only)         4+   ████████████  <- valuable
                                          pages buried here
Healthy: the bulk of pages — and all the important ones — sit at depth 1–3, with only minor pages deeper. Unhealthy: a heavy tail at depth 4+ that includes pages you actually want to rank.

Don't flatten everything: A long deep tail is only a problem if it contains pages that matter. Forcing every page shallow bloats navigation and dilutes links — the goal is shallow *important* pages, not uniform depth. Minor pages are allowed to sit deep.

Depth and authority are related, not the same

Depth is a useful proxy for internal authority, and treating them as identical will mislead you in both directions.

Depth counts hops on the shortest path. It does not care whether that path runs through a page with a hundred outbound links or one with five, or whether the linking page has any authority to pass. Two pages at depth 3 can be in completely different positions.

  • Shallow but starved — a page linked once from a sitewide footer is depth 1 and receives almost nothing, because the footer divides what it passes across every link on it.
  • Deep but well-supported — a page at depth 4 with several contextual links from strong, relevant articles can hold more internal authority than a depth-2 page linked only from a nav.
  • Shallow and strong — the intended case: a short path through pages that themselves have authority and few competing outbound links.

Which to trust When depth and authority disagree, authority is the better guide — it accounts for what the linking pages actually hold. Depth is the faster diagnostic and needs no modelling, which is why it is worth measuring first and acting on second.

This is also why "flatten everything" fails as a strategy. Adding every page to the navigation makes them all depth 1 while dividing the homepage's authority across hundreds of links — the arithmetic in link equity. The depth number improves and the pages are no better off.

Does depth still matter, given sitemaps?

It is a fair challenge. Search engines are far better at discovery than they were, sitemaps list every URL regardless of linking, and plenty of deep pages rank perfectly well. So why care about click-distance at all?

Because a sitemap solves discovery and nothing else. It tells a search engine a URL exists. It does not pass authority, carries no anchor text, and gives no signal about the page's importance relative to everything else you publish. Those come from links, and links are what depth measures.

  • Discovery — largely solved by sitemaps. A deep page will usually get found.
  • Crawl frequency — not solved. Sitemap presence does not make a page crawled often; link prominence does, which is why deep pages have stale cached versions.
  • Authority — not solved at all. A sitemap entry passes none. This is the gap that matters most.
  • Perceived importance — not solved. How you link a page is how you tell a search engine what you think of it, and a sitemap treats every URL identically.

The honest version Depth is a proxy, and a coarse one. It matters because it correlates with authority and crawl frequency, not because there is a rule against depth 5. A deep page with strong contextual links can outperform a shallow one that nothing meaningful points at — so treat a bad depth distribution as a symptom worth investigating, not a violation to correct.

The reason to measure it first anyway is cost. Depth needs only a crawl, and it reliably points at the parts of the site where the linking is weakest. Authority modelling gives the better answer; depth gives a decent answer immediately.

How it connects to the rest of structure

Depth isn't a standalone metric — it's downstream of your linking. Deep pages are usually deep because they're only reachable through pagination, deep filters, or a single buried nav link. Fixing depth means adding contextual links from shallow, authoritative pages, strengthening hubs, and wiring pages into clusters that pull whole topics up at once.

RankForge maps your full depth distribution, flags valuable pages stranded in the deep tail, and recommends the specific links that would flatten them — feeding the same dimension of the Structural Health Score. For the hands-on process, read how to improve crawl depth.

FAQ

What is a good crawl depth?

Keep important pages within about three clicks of the homepage or another strong entry point. Beyond depth three, crawl frequency and internal authority fall off sharply. The exact number matters less than the principle: priority pages shallow, only minor pages deep.

Is crawl depth the same as URL depth?

No. Crawl depth is click-distance through internal links; URL depth is how many folders are in the address. A page with a deeply-nested URL can be shallow if the homepage links to it directly, and a page with a flat URL can be deep if it's only reachable through many hops.

Why do my deep pages rank poorly?

Because depth costs them on two fronts at once: they're crawled less often (so updates are indexed late) and they receive only a heavily-decayed share of internal authority. Shortening the path with contextual links from shallow, authoritative pages usually helps more than editing the page.

Does crawl depth affect crawl budget?

Indirectly, and the direction is often misread. Search engines allocate more crawling to pages that look important, and link prominence is a large part of how importance is judged — so deep pages get recrawled less. The waste usually comes from the opposite side: crawl budget spent walking long pagination chains to reach content that could have been two clicks away.

How do I reduce crawl depth?

Add contextual links from shallow, authoritative pages, and build topic hubs that link directly to every page in a cluster instead of relying on sequential pagination. Adding everything to the site navigation lowers the number without helping, because the homepage's authority is then divided across hundreds of links.

What causes deep pages in the first place?

Pagination, overwhelmingly. A post drifts from page 1 of an archive to page 9 as newer content is published, gaining depth without anything about it changing. After that: deep filters and facets, sections hanging off a single navigation link, and nested documentation.

Should every page be within three clicks?

No — the goal is shallow important pages, not uniform depth. A long tail of minor pages at depth 5 is fine and normal. What matters is whether anything you actually want to rank is stranded out there, which is a question about which pages are deep rather than how many.

Do two SEO tools ever disagree on crawl depth?

Routinely, and usually for a legitimate reason. Depth depends on which links the crawler saw, so JavaScript rendering mode, whether nofollowed internal links are counted, how redirect chains are collapsed, and where the crawl budget ran out all change the answer. Compare like for like before treating a change in the number as a change in your site.

Is a flat site architecture better?

Flatter is usually better up to a point, and past that it backfires. Cramming every page into the navigation makes everything depth 1 while splitting the homepage's authority across hundreds of links, so each page receives less than it did before. The useful shape is a small number of strong hubs, each linking to a focused set of pages.

Sources

Put this into practice

Run the Crawl Depth Checker to see this on your own site, or run the full structural audit for the complete picture — both free, no account required.