NEW CASE
Antana

We create digital solutions that work for businesses


Give us a call +38 (066) 35-14-529

Let's take the first step towards your website — write to us

Close
Tools

Page Indexability Checker

Why a page is not getting into Google.

We check everything that can keep a page out of search: robots.txt rules for Googlebot, the robots meta tag, the X-Robots-Tag header, the status code, the redirect chain and the canonical.

Why your page is not in Google

Seven causes, most common first. The top three account for the majority of cases.

A noindex meta tag The page code tells search engines not to show it. Often left over from development, when the site was closed off and never reopened.
Blocked in robots.txt A Disallow rule keeps the crawler out. An over-broad rule can accidentally shut off an entire section.
Canonical points elsewhere The canonical tag tells Google another URL is the main version. This page then stays out of the index even if everything else is fine.
The page returns an error A 404 or 500 status. The page may look fine to a visitor while the crawler sees an error and moves on.
Nothing links to the page Google finds pages by following links. With none pointing at it — not from the menu, not from the sitemap — the crawler never learns it exists.
A duplicate of another page The text nearly matches another page on the site. Google picks one version and leaves the rest out.
The site is brand new Nothing is broken — its turn simply has not come. First pages appear within a week or two; full indexing takes up to a month.

noindex and Disallow are not the same

They get confused constantly, and the difference matters.

noindex

Do not show it

The crawler visits the page, reads it and sees the instruction to keep it out of search. Links on the page still count.

  • Placed in the page code
  • For utility pages and filters
Disallow

Do not visit it

The crawler never opens the page. But if links point to it, the URL can still surface in results — with no description.

  • Placed in the robots.txt file
  • To save crawl budget

The classic trap is using both at once. Disallow keeps the crawler out, so it never sees the noindex — and the URL stays in results. To remove a page from search you need noindex and open access to it.

How long indexing takes

A new site first pages in one to two weeks, full indexing within a month
A new page on an established site — from a few hours to a week
After a fix once noindex is removed, submit the URL in Search Console — that can cut it to a day
When to worry if a month has passed and the page is still missing, the cause is technical — check the list above

FAQ

How is this different from the meta tag checker?
The “Meta tags and Open Graph” tool shows what is written in the page tags. This one answers a different question: will the page enter search at all. Here we additionally download robots.txt and apply its rules to your exact URL the way Googlebot does, read the X-Robots-Tag header that is invisible in the page source, and check the redirect chain.
The page is open but missing from Google. Why?
Being indexable and being indexed are different things. If nothing blocks it, the usual causes are: the page is new and the crawler has not reached it; no internal link points to it, so Google does not know it exists; the content is too thin or nearly duplicates another page — Search Console then reports “Crawled — currently not indexed”. Submit the URL through the inspection tool and add links to it from other pages.
What if the canonical points to a different URL?
It means you have told Google: “index that page, not this one”. Sometimes that is intended — for filter or UTM-tagged pages, for example. But if the canonical is there by mistake, the page will never appear in results no matter how many links point to it. The classic breakage after a botched SEO plugin setup is every page canonicalising to the homepage.
Telegram Viber Call us