How Search Engines Actually Crawl and Index a Site
Guide · SEO Tools
Crawling, indexing, and ranking are three genuinely separate stages, and a page can succeed at one and fail at the next — understanding the difference explains a lot of confusing "why isn't my page showing up in Google" situations that otherwise seem mysterious.
Crawling: finding the page exists
A search engine's crawler discovers a page's existence by following links from other pages it already knows about, or from a sitemap you've submitted — a page with no links pointing to it and no sitemap entry may simply never be found at all, regardless of how good its content is.
Indexing: deciding to store the page
Once crawled, a search engine decides whether to actually add the page to its index — the database it searches through to answer queries. Pages can be crawled but not indexed for various reasons: perceived low quality, duplicate content, or a technical signal (like a noindex tag) telling the search engine not to include it.
Ranking: where it shows up, if at all
Only indexed pages are eligible to rank for any search query at all — being indexed doesn't guarantee ranking well, but not being indexed guarantees not showing up regardless of how relevant the content is. This is why "indexed but not ranking well" and "not indexed at all" are two completely different problems needing different fixes.
What actually helps at each stage
A sitemap helps crawling, by directly telling search engines what pages exist rather than relying purely on discovering links. A robots.txt file controls what crawlers are allowed to access at all. Quality, unique content and reasonable page structure help with indexing. Backlinks, content relevance, and page experience factors affect ranking. Treating all three as one undifferentiated "SEO" problem instead of diagnosing which stage is actually failing wastes effort on the wrong fix.