← all articles

What is crawl depth, and why it matters

Crawl depth is the number of clicks it takes to get from your homepage to a given page, following links the way a search engine crawler would. A page linked straight from the homepage sits at depth 1. A page linked from that page sits at depth 2. And so on down the tree.

I care about this because it quietly decides which pages get found, how often they get revisited, and which stay invisible for months. If some pages never show up in search, crawl depth is one of the first things I check. It is also cheap to fix, because it mostly comes down to how you link your own pages.

what it is

Crawl depth (also called click depth) measures distance, but distance in links, not in URL folders. The homepage is depth 0. Anything the homepage links to directly is depth 1. A blog post linked only from page 14 of your blog archive might be depth 15, even if its URL is just /blog/my-post.

A URL like /shop/shoes/running/mens/trail/ looks deep because of the folders, but if your main menu links straight to it, its crawl depth is 1. Folder structure and crawl depth are two different things. Crawlers follow links, not folder names.

Two related terms are worth knowing:

  • click depth: the same idea as crawl depth, usually used when you are talking about human visitors rather than bots
  • orphan page: a page with zero internal links pointing to it, so effectively infinite depth. a crawler can only find it through a sitemap or an external link, if at all

A rough mental model: the shallower the page, the more important your site is signalling it to be. Pages near the homepage get more internal links, more crawler attention, and more of whatever authority your homepage has built up.

how it works

Search engine crawlers start from known URLs. For a typical site that means the homepage, anything in your XML sitemap, and any URLs they have seen before. From each page they pull out the links, add the new URLs to a queue, and fetch them later. Each hop adds one level of depth.

Crawlers do not have unlimited time on your site. Google calls the budget it is willing to spend on a site the crawl budget, and its own documentation, Managing crawl budget for large sites, explains how crawl demand and server capacity limit what gets fetched. That page is written for very large sites, and Google says most small sites do not need to worry about budget as such. But the underlying behaviour still applies at any size: URLs that are easy to reach and clearly important get crawled first and most often.

So a deeper page tends to have these disadvantages:

  • it is discovered later, because the crawler has to get through several other pages first
  • it is revisited less often, so updates take longer to be noticed
  • it receives fewer internal links, and usually less link equity passed down from stronger pages
  • it is more likely to be skipped when the crawler decides it has seen enough

Your sitemap helps with discovery. The sitemaps protocol is simply a list of URLs you want crawlers to know about. But a sitemap is a hint, not a promise, and it does not fix the signal problem. A page that sits at depth 9 and is listed in a sitemap is still telling Google “this is not very important to me.”

You can measure crawl depth yourself. Screaming Frog’s SEO Spider (free up to 500 URLs; paid licences are listed on their site) reports a “Crawl Depth” column for every URL it finds. Sitebulb and Ahrefs Site Audit show the same thing. Run a crawl, sort by depth, and look at the distribution. A healthy small site usually has most of its important pages at depth 1 to 3.

My piece on reading a crawl report without panicking walks through which warnings deserve attention and which are noise.

why it matters

1. discovery and indexing

If a crawler cannot reach a page quickly, it may not be indexed quickly. I have seen new posts sit in “Discovered, currently not indexed” in Google Search Console for weeks, and the pattern is almost always the same: the page is buried, linked from one tag archive and nowhere else. Moving a link to it onto a high-traffic page often changes the picture within days.

When a page that used to be indexed drops out, depth is one of the suspects. I cover the wider checklist in when a page vanishes from the index.

2. how internal authority flows

Your homepage usually collects the most external links. Every internal link passes some of that value onward, and each extra hop dilutes it. Pages at depth 1 and 2 inherit the most. Pages at depth 6 inherit very little. If your money page or your best explainer is five clicks down, you are handing your strongest asset the weakest position.

3. freshness

Crawlers revisit shallow, popular pages more often. If you update a deep page, whether that is a price, a stock status or a corrected fact, the change can take a long time to be seen.

4. users give up too

Crawl depth is partly a proxy for a human problem. A visitor who needs six clicks to find what they came for will often leave. Fixing the structure for bots tends to help real people as well, which is the point Google makes in its SEO starter guide: build the site so people and crawlers can understand how pages relate.

where depth gets out of control

Most sites do not get deep on purpose. It happens by accumulation. The usual causes I run into:

  • pagination: blog archives and category pages where old posts only appear on page 12 or beyond
  • faceted navigation: filters that generate near-endless combinations of URLs, which both inflate depth and waste crawling. I wrote about this in faceted navigation and the pages that multiply
  • thin or duplicate pages: every low-value URL competes for the same crawl attention, a theme covered in index bloat explained
  • site migrations: old structures get wrapped in new ones and pages end up several levels lower than before

common misconceptions

“crawl depth is the same as URL depth”

It is not. As covered above, a page with four folders in its URL can be one click from the homepage, and a page with a flat URL can be buried. Check the link path, not the slashes.

“everything must be within three clicks”

You will see “keep everything within three clicks” repeated as a rule. It is a decent rule of thumb for small sites, but it is not a Google requirement. Large sites with hundreds of thousands of pages cannot put everything at depth 3. What matters is that your important pages are shallow and well linked. Low-priority pages can sit deeper without harm.

“a sitemap solves it”

A sitemap helps crawlers find URLs. It does not tell them a page matters, and it does not pass any link value. I treat the sitemap as a safety net and internal linking as the real structure. Likewise, blocking things in robots.txt (see Google’s robots.txt introduction) controls crawling, but it does not reduce depth for the pages you do want found.

“only huge sites need to care”

Google is right that crawl budget is mostly a large-site concern. But depth is not only about budget. On a 200-page site, a page buried under five layers still gets weaker internal signals and may still be slow to index. The effect is smaller, but the fix is quick, so I do not skip it.

a practical way to check your own site

Here is the routine I use on any site that has been running a while.

  1. Crawl the site with Screaming Frog, Sitebulb or similar and export the list of URLs with their depth.
  2. Take your 20 most important pages (the ones that earn money or that you want to rank) and read off their depth. Anything beyond 3 gets a note.
  3. Look for orphans: URLs in your sitemap or analytics that the crawl did not reach through links.
  4. Add contextual links from strong, shallow pages to the buried ones. A mention inside a relevant paragraph beats a footer link.
  5. Add a few hub pages, for example a clear “start here” page for each topic, so related articles share a short path.
  6. Re-crawl in a couple of weeks and compare.

If you are choosing what to prioritise on a young site, it helps to think about topic first. Choosing keywords for a site with no authority covers how I decide which pages deserve the best positions before I touch the link structure.

For readers who like to run crawl exports through AI summarisation to spot patterns, my sister site AI Tool Gazette covers that kind of tooling. I would still read the raw depth column yourself once; it takes five minutes.

where to go from here

If crawl depth was new to you, these are the next things I would read, in roughly this order:

You can also browse everything else on the blog index.

Written by Xavier Fok

disclosure: this article may contain affiliate links. if you buy through them we may earn a commission at no extra cost to you. verdicts are independent of payouts. last reviewed by Xavier Fok on 2026-10-09.

for SEOs
Tracking rankings or scraping SERPs at scale?

Rank checkers and SERP crawlers get blocked and geo-skewed fast on datacenter IPs. Singapore Mobile Proxy runs real 4G/5G mobile IPs that search engines still trust, so your position data stays clean.

see plans →
read on
More from The SEO Desk

Technical SEO, link building, content and SERP strategy, and tool reviews for people who ship growth.

browse all articles →