What is indexing and why a page is missing from Google - Zephyra Studio
Indexing means Google has found a page, read it, and added it to the database it later draws results from. If a page is not indexed, it cannot appear for any query, however good the content is. So the first question about any new page is exactly that: is it indexed at all. The check is quick, and the reasons a page is missing are usually technical and easy to fix.
How to check whether a page is indexed
The quickest check is URL Inspection in Google Search Console: enter the page address and read the status. If it says the URL is on Google, the page is indexed.
A second, rougher check is a site:yourdomain search. If the page does not appear in those results, it is probably not indexed, although a site: search is not always fully accurate and serves only as an indicator.
The most common reasons a page is not indexed
Google lists several statuses in Search Console. Discovered, currently not indexed means the URL was found but not yet processed. Crawled, currently not indexed means it was read but not added, most often because of thin or duplicate content.
A third common status is Excluded by noindex tag, when a page tells Google not to index it. A fourth is Blocked by robots.txt, when access is forbidden. Each of these statuses has a concrete fix, and Search Console names it precisely.
Tags that block indexing
noindex is a tag in the head of a page: <meta name="robots" content="noindex">. It most often stays behind by accident from the development phase, when it made sense for a test version not to appear in Google. After launch it should be removed.
robots.txt is the other file, often misunderstood. A block in robots.txt stops Google from reading a page, but does not remove it from the index if it is already there. To remove a page from the index, use noindex, not robots.txt.
A practical tip: the order of the check
The order of the check: first URL Inspection in Search Console, then, depending on the status, look at three things: whether the page has a noindex, whether it is blocked in robots.txt, and whether it is included in the XML sitemap.
Finally, check whether there are internal links to that page. A page reachable only from the sitemap, with no link from the site itself, struggles to get Google's attention. A link from the menu or from body text is the strongest route to indexing.
Related terms
For the wider picture, see also:
Source
Key takeaways
- Indexing means Google has read a page and added it to the database it draws results from.
- A page that is not indexed cannot appear for any query.
- URL Inspection in Search Console is the quickest status check.
- noindex removes a page from the index, robots.txt only stops reading it, and internal links speed up indexing.
Conclusion
If a page is not indexed, everything else is irrelevant until that is fixed. Our technical SEO work includes an index check as the first step, because without it there is no point writing new content or rewriting titles.
Frequently asked questions
Indexing means Google has found a page, read it, and added it to the database it later draws results from. If a page is not indexed, it cannot appear for any query, however good the content is. So the first question about any new page is exactly that: is it indexed at all. The check is quick, and the reasons a page is missing are usually technical and easy to fix.