Crawling is Googlebot visiting a page, and the index is the database the visited page is recorded in. These are 2 separate steps: a page can be crawled and not indexed. If it is not in the index it cannot appear in search, however good it is.
Googlebot most often arrives through links. It follows connections from pages it already knows and finds new ones that way. A page with no link pointing to it - internal or external - can stay undiscovered for a long time.
The second route is the sitemap. A submitted sitemap speeds up discovery, especially on new sites and for pages deep in the structure.
The third is manual submission of an address through Search Console. Useful for a single page that urgently needs indexing, but no substitute for a sound link structure.
The most common reason is quality. Google does not index everything it crawls - thin pages, near-empty categories and duplicates are often skipped. In the report this appears as "crawled, currently not indexed".
The second reason is your own directives. If a page carries noindex or a canonical pointing elsewhere, you explicitly said it should not be indexed.
The third is crawl budget. On large sites the crawler does not visit everything every time, so pages with no internal links and no visits get crawled rarely.
On small sites crawl budget is practically never the issue. If a page is not indexed, the cause is almost always content quality or a misplaced directive.
The quickest check is the URL inspection tool in Search Console. Paste an address and you get an answer on whether the page is indexed, when it was last crawled and what the crawler saw.
For a whole-site view there is the indexing report. It groups pages by reason for exclusion, so it is immediately clear whether the problem is directives, duplicates or quality.
The site: operator in search gives a rough picture but is unreliable for exact counts. For decisions, use Search Console.
First check the directives - noindex, canonical and robots.txt. These are minute-long fixes that explain a large share of cases.
Then check whether any internal link leads to the page. A page you cannot click through to from navigation or body text looks unimportant to both the crawler and the user.
If the technical side is clean, the problem is content. A page of two hundred words repeating what is written elsewhere on the site rarely enters the index - and no technical fix helps there, only better content.
After making changes, request a recrawl through Search Console. Indexing can take anywhere from a few days to a few weeks.
This is the message that most confuses site owners. It does not mean something is broken - it means the crawler visited the address but judged it not worth recording in the database.
The reason is almost always a thin page: a few sentences, content repeated from other sub-pages, or an empty category waiting to be filled. Such pages accumulate on online stores and on sites with many filters.
The solution is not technical. Either the page gains content that genuinely answers a question, or it is merged with a related page, or it is explicitly excluded from the index so it does not dilute the rest of the site.
It is also important not to panic over new pages. For the first few weeks after publishing, that status is normal, because search still has to work out where the page belongs.
Last updated: 17 August 2026
Send the address and we will tell you where you stand - no obligation.
Send an enquiry