Every page that ranks on Google had to go through two separate steps first: crawling and indexing. Confusing the two is one of the most common reasons site owners misdiagnose why their content “isn’t ranking” — when the real problem is it was never indexed in the first place.

Table of Contents

Crawling vs. Indexing vs. Ranking

These are three distinct stages, and a page must pass through all of them in order:

A page can be crawled but not indexed (if it’s low quality or blocked). It can be indexed but rank poorly. Getting indexed is necessary but not sufficient — see our guide on whether you need to submit your site to Google for the discovery side of this process.

How Googlebot Crawls a Website

Googlebot works through a queue of URLs, prioritized by factors like site authority, update frequency, and crawl budget. On each visit it:

  1. Requests the page and downloads its HTML (and increasingly, renders JavaScript)
  2. Extracts all links on the page to add to its crawl queue
  3. Checks your robots.txt file for any crawling restrictions
  4. Reads your XML sitemap for a prioritized list of URLs

Mobile-First Indexing

Since 2020, Google predominantly uses the mobile version of your site’s content for indexing and ranking, using a mobile crawler by default. In practice this means: if content, links, or structured data are missing or hidden on your mobile layout compared to desktop, Google may simply never see them. Checking that your mobile and desktop versions carry equivalent content is worth doing even if most of your traffic happens to be desktop.

What Determines Crawl Budget

Larger or lower-authority sites need to think about crawl budget — the number of pages Google is willing to crawl within a given timeframe. Key factors:

For most small to mid-sized sites, crawl budget isn’t a binding constraint — it becomes relevant mainly once a site reaches the tens of thousands of pages, or has significant technical waste like faceted navigation generating near-duplicate URLs.

Common Reasons Pages Don’t Get Indexed

You can check indexing status for any URL using Search Console’s URL Inspection tool, or the Pages report under the Indexing section. For the full technical reference, see Google’s crawling and indexing documentation.

Reading the Pages Report in Search Console

The Pages report groups your URLs into indexed and non-indexed buckets, each with a specific reason. A few of the most common non-indexed reasons and what they typically mean:

Frequently Asked Questions


Crawling and indexing is the first stage of a bigger process — see our SEO Basics guide for the complete picture.

Want a technical crawl audit of your own site? Web Solution Zone’s SEO team can identify exactly what’s blocking your pages from being indexed.

Leave a Reply

Your email address will not be published. Required fields are marked *