NEW CASE
Antana

We create digital solutions that work for businesses


Give us a call +38 (066) 35-14-529

Let's take the first step towards your website — write to us

Close
August 14, 2026 9 min read

Why a Website Isn't Being Indexed by Google: Reasons and a Step-by-Step Check

SEO
Why a Website Isn't Being Indexed by Google: Reasons and a Step-by-Step Check

A page can be published, open correctly in a browser, and contain useful copy while remaining absent from Google. This does not automatically mean a penalty or “bad SEO”. In most cases, the cause sits at one of three levels: Google has not discovered the URL, cannot crawl it, or has decided not to add it to the index.

The workflow below avoids random fixes. Start with the exact URL in Google Search Console, confirm access and directives, and only then inspect canonical signals, duplication, internal links, and the page’s independent value.

Indexing and ranking are different problems

Indexing means Google has stored a page in its index and may show it in Search. Ranking determines the queries and positions for which it appears. If a URL is not indexed, rewriting the title or adding keywords will not resolve the underlying issue.

Situation What it means First action
URL is unknown to Google The page has not been discovered Add an internal link and check the sitemap
URL is known but not crawled Crawling is delayed or technically blocked Check the server, robots.txt, and link structure
Crawled but not indexed Google saw the page but did not add it Review quality, duplicates, canonical, and soft 404 signals
Indexed but receiving no traffic The issue concerns relevance and ranking Analyse queries, content, CTR, and competition

A ten-minute diagnostic

1. Inspect the exact address

Open Google Search Console, paste the full URL into the top bar, and review its status. The URL Inspection tool reports what Google knows about the indexed version and can test whether the live URL is accessible and potentially indexable.

Use the final address with the correct protocol, hostname, language folder, letter case, and trailing slash. Different variants can be treated as separate URLs.

2. Test the live URL

The indexed report may describe an older version. A live test shows what Googlebot can access now, including the response, crawl permission, directives, and processed HTML.

3. Compare canonical URLs

Compare the user-declared canonical with the canonical selected by Google. If they differ, Google considers another address the primary version. Google recommends aligning canonical declarations, internal links, and sitemap entries rather than sending conflicting signals.

4. Review the Page indexing report

The Page indexing report groups known URLs by the reasons they were not indexed. Focus on whether important service, category, product, or article pages appear in an exclusion group, not merely on the total number of excluded URLs.

5. Request indexing only after the fix

The request button does not repair a page. Resolve the cause, repeat the live test, and then submit the URL. Google says recrawling can take from several days to several weeks and does not guarantee inclusion in the index.

Common reasons a page is not indexed

Google has not discovered the URL

A new page without internal links is effectively orphaned. Add it to a relevant category, link from pages Google already visits, and include the canonical URL in the XML sitemap. A sitemap helps discovery, but it does not guarantee crawling or indexing.

robots.txt blocks crawling

An accidental Disallow for a directory, language version, parameter, or the entire site can stop crawling. However, robots.txt is not a mechanism for removing a web page from Google. A blocked URL may still appear without a useful snippet if other pages link to it.

A noindex directive remains

noindex in a robots meta tag or X-Robots-Tag header explicitly prevents indexing. It often survives a staging migration, template change, or SEO plugin setting. Google must crawl the URL to read the directive, so blocking the same page in robots.txt creates a contradictory setup.

The server returns the wrong response

An indexable page should reliably return 200 OK. Redirect chains or loops, 4044105xx errors, unstable security rules, or Googlebot blocking interfere with crawling. Test from outside the signed-in session and inspect server logs where possible.

Canonical points elsewhere

If canonical points to the homepage, a category, another language version, or a parameter-free URL, the current page may not be indexed separately. A unique page generally needs a self-referencing canonical, while internal links should use that same preferred address.

Google sees a duplicate or soft 404

Filters, sorting, parameters, tag archives, and near-identical product cards can create large duplicate sets. A soft 404 returns 200 but looks empty, erroneous, or lacking independent value. Not every technical URL combination should be indexable.

Primary content depends on JavaScript or differs on mobile

Google can process JavaScript, but rendering adds another stage and more failure points. Do not require a click to reveal primary copy, links, or product data. Under mobile-first indexing, the mobile version is the basis for indexing, so important content, robots directives, and structured data should be consistent.

The page has no distinct purpose

Technical accessibility does not oblige Google to index a URL. A page that repeats another resource, contains only generic paragraphs, or does not satisfy a distinct intent may not qualify as a useful separate result. Add a complete answer, practical evidence, clear authorship, sources, and relevant internal links.

Team analysing the causes of page indexing problems

A structured review of access, directives, and canonical signals reveals the actual cause of an indexing problem.

How to interpret common Search Console statuses

Status Likely explanation Action
Discovered — currently not indexed Google knows the URL but has delayed crawling Improve internal links, sitemap signals, speed, and server stability
Crawled — currently not indexed Google crawled the page but did not select it Review value, duplicates, canonical, and soft 404 signals
Excluded by noindex tag An indexing prohibition was found Remove noindex if the page belongs in Search, then retest
Blocked by robots.txt Googlebot cannot crawl the URL Correct the rule if the block was not intentional
Duplicate, Google chose different canonical Another primary version was selected Align canonical, redirects, sitemap, and internal links
Page with redirect The URL redirects elsewhere Confirm the redirect and remove the old URL from the sitemap
Soft 404 The content resembles an empty or error page Add genuine value or return an appropriate 404/410

Interpret every status according to the URL’s purpose. Excluding a cart, sign-in page, or technical filter is often correct. The same outcome for a service or category page requires investigation.

Step-by-step recovery plan

  1. Choose the preferred canonical URL. Define the single address intended for Search.
  2. Verify the HTTP response. The page should consistently return 200 OK without loops or blocks.
  3. Check robots.txt, robots meta, and X-Robots-Tag. Remove unintended restrictions.
  4. Align canonical, hreflang, and sitemap. They should not name competing versions.
  5. Inspect the HTML available to Googlebot. Primary content, headings, and links must be accessible without interaction.
  6. Add internal links. Connect the URL to its category and relevant supporting pages.
  7. Strengthen the page. Satisfy a specific intent fully and remove unnecessary duplication.
  8. Update the sitemap and repeat the live test. Request indexing after a successful result.
  9. Monitor rather than resubmit every day. Watch the URL and its issue group in the report.

Diagnostic mistakes to avoid

  • Relying only on the site: operator as a complete index report.
  • Requesting indexing after every minor edit.
  • Blocking a URL in robots.txt while expecting Google to read its noindex.
  • Listing redirects, errors, duplicates, and non-canonical URLs in the sitemap.
  • Creating multiple pages for near-identical keywords without distinct value.
  • Changing URLs, canonicals, and architecture together without a redirect map.
  • Rewriting titles while the page is not crawlable.

Checklist before requesting indexing

  • The URL works without authentication and returns 200 OK.
  • robots.txt permits crawling.
  • No noindex exists in HTML or HTTP headers.
  • Canonical names the intended indexable version.
  • The URL appears in the current sitemap without a redirect.
  • Relevant internal pages link to it.
  • Primary content is available to Googlebot on mobile.
  • The page is not a duplicate and offers independent value.
  • The title, H1, and content answer one search intent.
  • The Search Console live test reports no critical block.

Frequently asked questions

How long does Google take to index a new page?

There is no fixed deadline. Google gives a range of several days to several weeks for recrawling. Accessibility, internal links, crawl patterns, and page quality all influence the outcome.

Does a sitemap guarantee indexing?

No. It helps Google discover preferred URLs but does not force crawling or indexing. Include only canonical 200 pages that genuinely belong in Search.

Should I request indexing every day?

No. Repeated requests do not replace a fix or guarantee faster processing. Submit after a material change and monitor the report.

Why did Google choose a different canonical?

Signals may conflict: the canonical declaration identifies one URL, the sitemap another, and internal links a third. Pages may also be similar enough for Google to group them as duplicates.

Can an indexed page still receive no impressions?

Yes. Indexing only makes a URL eligible to compete. Impressions still depend on relevance, quality, clear structure, and competitiveness.

What to do next

When one important page is missing, work through the checklist from top to bottom. If the issue affects categories, localized versions, or hundreds of URLs, perform a technical SEO audit of templates, indexing rules, sitemap files, canonicals, and site architecture.

Explore more guidance in the SEO promotion section. For a systematic diagnosis, see BB STUDIO SEO services or send us your website address for an initial assessment.


Prepared by the BB STUDIO team. This article draws on our technical website review practice and official Google Search Central and Google Search Console documentation.

Rate this article
It helps us write better content
Be the first to rate 5.0 of 5 0 votes

Recommended reading

Let’s create something amazing together

Become a clientBecome a client
Telegram Viber Call us