Definition
Indexing is when a search engine processes a crawled page and stores it in its index, making it eligible to appear in search results.
Indexing explained
After crawling a page, a search engine analyzes its content, works out what it is about, chooses a canonical version among duplicates, and decides whether to add it to the index. Only indexed pages can appear in results, including as sources for AI Overviews.
Crawled does not mean indexed. Search engines skip pages they consider duplicates, low quality or not useful enough, and pages that carry a noindex rule. Google Search Console's page indexing report groups excluded pages by reason, with statuses such as “Crawled - currently not indexed”, “Discovered - currently not indexed” and “Duplicate without user-selected canonical”.
To improve indexing, make sure important pages are linked internally and listed in the sitemap, remove accidental noindex tags, consolidate duplicates with canonical tags or redirects, and improve thin pages or remove them. Requesting indexing in Search Console can speed up individual pages but doesn't override quality decisions.
Example
Search Console shows many of your product pages as “Crawled - currently not indexed”. They share near-identical descriptions from the manufacturer. Rewriting descriptions for the most important products, with your own sizing notes and photos, gives Google a reason to index them.
Why it matters
Indexing is the gate between publishing a page and having any chance of search traffic or AI citations from it.
Related service
Technical SEO
Crawlability, indexing, structured data and rendering fixes so search engines and AI crawlers can read your site.
Technical SEO servicesPublished by Vidern, founded and led by Malhar Shah. Updated .
Find out what's holding your site back
Get a free SEO audit with prioritized fixes, delivered within 48 hours.
Related terms
- CrawlingCrawling is the process by which search engine and AI bots discover web pages by following links and sitemaps, then download their content for processing.
- NoindexNoindex is a rule, set in a robots meta tag or an X-Robots-Tag HTTP header, that tells search engines not to include a page in their results.
- Canonical tagA canonical tag is an HTML link element that tells search engines which URL is the preferred version of a page when similar or duplicate versions exist.
- Google Search ConsoleGoogle Search Console is Google's free tool showing how a site performs in Search, which pages are indexed, and technical problems Google finds.
- Duplicate contentDuplicate content is substantially the same content appearing at more than one URL, on the same site or across different sites.