What causes crawled currently not indexed is rarely one thing: the status means Google fetched the URL and did not add it to the index at that point in time, and the reasons range from near-duplication and low-value URL patterns to brand-new pages, site-level signals and plain selectivity. It is a state rather than a penalty, and in any SEO South Africa audit the first job is establishing which cause applies before changing a word of content.

Google's own documentation is direct about the selectivity. The Search Console page indexing report says the page "may or may not be indexed in the future" and tells site owners "don't expect every URL on your site to be indexed". Google's documentation on how Search works goes further: "not every page that Google processes will be indexed." A site carrying some URLs in this state is not automatically broken — which is why diagnosis comes before the work in the guide on improving website indexation.

Quick Answer

What causes crawled currently not indexed is a set of separate causes, not a single verdict: near-duplication where Google selected a different canonical, a page that falls short on its own merits, site-level quality signals, filter and parameter URLs, pages nothing links to, and ordinary selectivity — Google states that not every page it processes gets indexed. The status is a point-in-time state, and URLs move out of it. Establish the cause with URL Inspection first: near-duplication is resolved by canonicalisation, a filter URL by crawl control, and a site-level pattern by work nowhere near the affected page.

Pages sitting unindexed in Search Console?

Send us your Search Console Page Indexing export and we will tell you which of the causes below is actually behind your affected URLs.

Get a Free Indexation Review

What Causes Crawled Currently Not Indexed: Six Explanations

"Crawled — currently not indexed" is one Search Console label covering at least six distinct situations, which is why any single prescribed fix fails most of the time. Google's documentation supports several of them independently: canonical selection during indexing, low crawl demand, the useless URL spaces produced by faceted navigation, the weeks it can take Google to notice a new site, and the plain statement that indexing is not assured. The table maps each cause to the evidence identifying it and the matching fix.

CauseWhat it looks likeEvidence that identifies itVerdict: the matching fix
Near-duplicate of an indexed pageAnother URL carries substantially the same primary content, and Google grouped them during indexing.URL Inspection's "Google-selected canonical" names a different URL from the one you inspected.Consolidate the variants into the preferred URL. Rewriting the excluded page achieves nothing while it stays a duplicate.
The page's own valueNo twin exists, the page is linked, and it is thinner than what already ranks for its query.Google-selected canonical matches the URL; a verbatim-sentence search finds no near-copy; the page loses a side-by-side comparison.Rewrite it to answer the query more completely, merge it, or retire it. The only cause a rewrite resolves.
Site-level quality signalsThe status spreads across every template — blog, product, category, service — rather than concentrating in one.Grouped by template, no single type dominates, and the affected share is large relative to what is indexed.Site-wide content audit. Google's core update guidance asks you to assess "your site overall (not just individual pages)".
Filter, parameter and faceted URLsAffected URLs carry sort, filter, colour or pagination parameters generated by the storefront, not authored as pages.The URLs contain query strings, and the count scales with catalogue size rather than publishing activity.Crawl control — robots.txt, canonicals, nofollowed filter links. Often the status is the correct outcome and needs no fix.
New page, new site, or nothing linking to itThe URL is recent, or the domain is, and the sitemap is its only route in.Published within weeks; a site crawl shows zero inbound internal links; Discovery shows sitemap only.Add contextual internal links from related pages, then allow time. Google says it can take a few weeks to notice a new site.
Ordinary selectivityA handful of isolated URLs on a site indexing normally, with nothing wrong with the pages.No template pattern, no duplication, links present, and the page holds up against what ranks.Verify and monitor. Google's report guidance says not to expect every URL to be indexed, and URLs move in and out of this state.

What Google Has Actually Said About This Status

Search Engine Journal reported Gary Illyes, a Google Search Advocate, at SERP Conf 2024 calling the label deliberately broad: "ideally we would break up that category into more granular chunks, but it's super hard because of how the data internally exists. It can be a bunch of things." He named duplicate elimination, a site error serving "the same exact page to every single URL on the site", and a change in "our perception of the site" — closing with "there could be many things". That is a reported conference remark, not documentation, but it points where the documentation points: the label is a bucket, not a diagnosis.

The Premise That Wastes the Most Time

Treating "Crawled — currently not indexed" as a verdict on content quality sends most sites straight to a rewrite, which is the right response to exactly one of the six causes. A near-duplicate needs consolidation, a filter URL needs crawl control, an unlinked page needs internal links, and a site-level pattern needs work on pages that are not in the affected list at all. Identify the cause and the fix follows; guess at it and the pages stay where they are.

Crawled vs Discovered: Two Different States in One Report

Google Search Console reports "Crawled — currently not indexed" and "Discovered — currently not indexed" as separate states with different meanings: the first means Google fetched the page and did not index it, while the second, in Google's own wording, means "the page was found by Google, but not crawled yet", typically because crawling it was expected to overload the site. The two share a report and almost nothing else.

StatusWhat Google has doneLast crawl in URL InspectionVerdict: where the fix lives
Discovered — currently not indexedFound the URL; has not fetched itEmptyCrawl capacity and crawl demand — server performance, internal links, less wasted crawling (see crawl budget)
Crawled — currently not indexedFetched the page; did not index it at that timePopulated with a dateOne of the six causes above — canonical selection, page value, URL pattern, linking, site-level signals, or nothing

The Last crawl field is the cheapest way to tell them apart, and it is worth reading before anything else. An empty Last crawl means a discovery and scheduling problem; a date means Google has the content and made an indexing decision on it. The crawl side is covered in the guide on how to improve crawlability.

The Diagnostic Sequence in Google Search Console

Diagnosis is a fixed sequence of checks that separates the six causes from one another using evidence available in Search Console and a site crawl, and it runs before any content or technical work begins. The order matters: checks that rule out whole categories come first, and the comparison against ranking competitors — the check that leads to rewriting — comes last, because it costs the most and applies to the fewest URLs.

Step 1: Export the affected sample, and treat it as a sample. Open Indexing → Pages, find "Crawled — currently not indexed" under the reasons list, and export. Google's documentation notes the example list "does not necessarily show all URLs with that issue, and is limited to 1,000 rows", so the export shows the shape of the problem rather than its full extent. That crawled not indexed Google Search Console export is your diagnostic input, not a re-indexing queue.

Step 2: Read the Google-selected canonical on three or four representative URLs. If it names a different URL, the cause is canonical selection — your page was grouped with another and the other one was chosen. Google's canonicalisation documentation is explicit that a declared canonical "is a hint, not a rule", so a self-referencing tag on your page does not settle the question.

Step 3: Check the Last crawl date and HTTP response. A recent crawl date and a 200 response confirm you are in the crawled state rather than the discovered one, and rule out access as the cause. An empty Last crawl means you are diagnosing the wrong status.

Step 4: Look at the URL shape. Sort the export by URL pattern. Query strings, sort and filter parameters, session identifiers and deep pagination point to the faceted navigation cause, which Google's guidance describes as crawlers accessing "a very large number of faceted navigation URLs" before determining they are useless. Those sitting unindexed is the system working; the guide on faceted navigation and SEO covers how to stop generating them.

Step 5: Group the remaining URLs by template. One template dominating points to a structural cause inside that template. An even spread across blog posts, products, categories and service pages points instead at a site-level pattern — the cause where individual page work has the weakest claim on your time.

Step 6: Count internal links to each affected URL. Crawl the site and filter for pages with zero inbound internal links. Google's SEO Starter Guide states that "the vast majority of the new pages Google finds every day are through links", which makes an unlinked page a discovery problem before it is a content problem. The process is in the guide on finding orphan pages.

Step 7: Only now, compare the page against what ranks. For URLs that survive steps 2 to 6 — self-canonical, crawled, cleanly structured, linked, not part of a site-wide pattern — open the top results for the target query and compare. A materially thinner or less specific page is the page-value cause, and a rewrite is the right answer for it.

Why the Order Is Not Negotiable

Running the content comparison first is how teams end up rewriting near-duplicates that were never going to be indexed under their own URL, and re-optimising filter URLs that should not exist. Reading the Google-selected canonical takes about thirty seconds per URL and eliminates the most common cause before any writing starts. Every step above rules a cause in or out, and none requires a tool beyond Search Console and a site crawler.

Not sure which cause your URLs fall under?

Share your Search Console export and we will run the diagnostic sequence and name the dominant cause before you spend a cent on content.

Request a Technical SEO Assessment

Matching the Fix to the Cause

Each of the six causes has one fix that resolves it and several that do nothing for it, which makes applying the wrong remedy the most expensive error in indexation work. The crawled currently not indexed fix you need depends entirely on what the diagnostic sequence returned.

When Google selected a different canonical

Consolidation is the fix, not rewriting. Decide which URL should survive, redirect or merge the others into it, and repoint the internal links. Google's documentation says a canonical declaration is a hint rather than a rule, so a self-referencing tag on a genuine near-copy does not force the outcome — the duplication itself has to go. Implementation is covered in the guide on handling duplicate content.

When the page falls short on its own merits

Rewrite the page to answer its target query more completely than the pages currently ranking, or fold it into a page that already does. Word count is not the mechanism; information the ranking pages do not carry is — local pricing, a decision table, a worked example. If the page cannot be improved into something worth indexing, retire it and drop it from the sitemap. Where the list of weak pages is long, our SEO service for South African sites can work through improve, merge or retire decisions on a flat monthly retainer, starting with a free audit.

When the affected URLs are filter and parameter variants

Stop generating the URLs rather than trying to get them indexed. Google's faceted navigation guidance points to blocking those paths in robots.txt, using URL fragments where a filtered view does not need its own URL, applying rel="canonical" to consolidate variants, and nofollowing filter links. Filter URLs sitting unindexed is a normal outcome; the work is reducing how many exist.

When nothing links to the page

Add three to five contextual internal links from topically related pages that already perform, placed inside body paragraphs rather than navigation or footers. Google's documentation on crawl budget states that when crawl demand is low, Google crawls a site less — and internal links are the demand signal you control. Then allow time: Google says it can take a few weeks to notice a new site.

When the pattern is site-wide

No published Google source gives a percentage above which a count becomes a site-level signal, so judge it by shape, not threshold: an even spread across every template, growing over months. The intervention is a content audit across the site — consolidating, improving and removing weak pages, as covered in content pruning and finding declining SEO pages, sequenced using prioritising SEO fixes. Google's core update guidance says changes can take a few days to several months to confirm.

What requesting indexing does and does not do

Requesting indexing in URL Inspection asks Google to recrawl sooner. Google's documentation states that "submitting a request does not guarantee that the page will appear in the Google Index", and there is a daily limit on submissions. It belongs at the end of the sequence, after the change is live and verified — requesting a recrawl of an unchanged page asks Google to look again at what it already assessed.

The Right Order of Operations

Pages crawled but not indexed need the cause identified, then the matching change, then a recrawl request — in that order. Diagnose which of the six causes applies, make the change that cause calls for, confirm it is live, and only then use URL Inspection to request indexing. Skipping to the re-indexing request is the most common reason Google Search Console crawled currently not indexed counts stay flat despite repeated resubmission.

The Patterns That Catch South African Sites

South African sites concentrate in three of the six causes for structural reasons rather than anything about local search behaviour, and knowing which three shortens most audits considerably.

Near-duplicate metro service pages. Agencies, installers and legal firms across Johannesburg, Cape Town, Durban and Pretoria routinely publish one page per metro with the city name swapped and little else changed. Those fall squarely into the canonical selection cause. Check the Google-selected canonical before touching the copy — if it names a sibling page, the answer is consolidation, or genuine localisation with different pricing context, suburbs and examples a reader in that city would recognise.

WooCommerce and Shopify filter URL spaces. SA stores running large catalogues generate colour, size, price-range and sort URLs by the thousand, and those inflate this status with no content problem behind them. The affected count scaling with catalogue size rather than publishing activity is the tell. Treat it as crawl control, not content.

Load-shedding, which produces a different status. Repeated server downtime during outage windows surfaces in Search Console as "Server error (5xx)", not as "Crawled — currently not indexed". If your report shows 5xx errors clustering at predictable times, the fix is hosting infrastructure with generator backup or cloud redundancy. Where both appear, treat them as two problems with two fixes.

Why South African Businesses Choose Growth Pulse Media

Growth Pulse Media was founded by Dirk van Greuning, who built and scaled a large South African ecommerce business before turning to agency work. That background is why indexation work here starts with Search Console evidence rather than a content proposal — thin product pages, metro service pages that cluster into each other, and blog archives with no impressions in twelve months are three different problems, and only one is solved by writing.

Technical SEO, content audits and indexation recovery are handled in-house, with a deliberately limited client load so every account gets senior attention. Businesses in Johannesburg and Gauteng can see the full scope on the SEO South Africa service page. The daily toolset is Search Console, Screaming Frog and Ahrefs, and the registered credentials include Shopify Partner and Omnisend Certified Partner status.

Who This Fix Is Not For

Sites that want the count to reach zero. Google's page indexing report says it outright: "Don't expect every URL on your site to be indexed." Filter URLs, paginated variants, thin tag archives and near-duplicates sitting unindexed is the system behaving correctly. If the brief is "get this number to zero", the brief is wrong, and chasing it burns budget on URLs that should stay out of the index.

Businesses that want the pages indexed without diagnosis or change. There is no technical switch that indexes a near-duplicate or a filter URL as it stands. Re-submitting sitemaps, requesting indexing and adding canonical tags are all useful — after the diagnostic sequence has said which cause you are dealing with. An agency promising indexation without identifying the cause first is guessing with your money.

Sites with a site-wide pattern and no appetite for a site-wide audit. When the status spreads evenly across every template, the work sits on pages that are not in the affected list — consolidating, improving and removing weak content across the site. That is a months-long programme, and Google's core update guidance says its systems can take several months to confirm the change. Page-level fixes will not substitute for it.

Publishers running mass generated content with no editorial layer. Practitioners consistently report elevated rates of this status on sites publishing high volumes of unreviewed AI output, though no platform documentation confirms a specific mechanism. What is documented is that Google's core update guidance asks whether the site overall delivers helpful, reliable, people-first content — which puts a volume-first content process upstream of the indexation symptom, not downstream of it.

Sitting on a backlog of unindexed URLs?

Send us the affected list and we will separate the URLs worth fixing from the ones that should stay out of the index — no obligation, and we get back to you within 24 hours.

Book a Free SEO Audit

Frequently Asked Questions

What does "crawled — currently not indexed" mean in Google Search Console?

It means Google fetched the page and did not add it to the index at that time. Google's page indexing report describes it as a page that "may or may not be indexed in the future", so it is a point-in-time state rather than a permanent judgement, and URLs move in and out of it. The status names an outcome, not a cause.

Is "crawled — currently not indexed" always a content quality problem?

No. Content value is one of several causes, alongside near-duplication where Google selected a different canonical, filter and parameter URLs, pages nothing links to, brand-new pages and sites, and ordinary selectivity. Google's documentation on how Search works states that "not every page that Google processes will be indexed." Rewriting a page affected for one of the other reasons will not change the outcome.

Is "crawled — currently not indexed" the same as a noindex tag?

No. A noindex directive is an explicit instruction in the page's meta tags or HTTP header, and Search Console reports it separately as excluded by a noindex tag. The "crawled — currently not indexed" status means no such instruction was present and the page still was not indexed. If both appear in your report, treat them as two lists with two causes.

How do I tell whether duplication or page value is the cause?

Open the URL in the URL Inspection tool and read the "Google-selected canonical" field. If it names a different URL, your page was grouped with another one and that other page was chosen, which makes the fix consolidation rather than rewriting. If the Google-selected canonical matches the URL you inspected and the page is linked and crawled, page value becomes the plausible cause and a comparison against the ranking pages is the next check.

Does requesting indexing in Google Search Console fix the status?

Requesting indexing asks Google to recrawl sooner, and Google's documentation states that "submitting a request does not guarantee that the page will appear in the Google Index." It does not change what Google finds when it arrives. If the page is still a near-duplicate, still a filter URL or still unlinked, the same decision follows. Make the change the diagnosis called for, verify it is live, then request indexing.

Are South African service-area pages more likely to get this status?

They are a common example of the canonical selection cause, because one page per metro with only the city name changed is exactly the near-duplicate pattern Google groups together during indexing. Check the Google-selected canonical on a few of them before assuming a quality problem. Where the pages cannot be genuinely differentiated with local pricing, suburbs and examples, consolidating to one strong page beats maintaining several that compete with each other.

Get the Right Fix for the Right Cause

Growth Pulse Media diagnoses "Crawled — currently not indexed" from your Search Console evidence first — canonical selection, URL patterns, internal linking, page value or a site-level pattern — and only then proposes work. Dirk and the team work directly in your Search Console and crawl data, using the same Screaming Frog, Ahrefs and Google Search Central toolset every day, across WooCommerce and Shopify catalogues as well as service sites.

No obligation — we respond within 24 hours.

Request Your Free Indexation Audit
Dirk van Greuning — Founder, Growth Pulse Media
Dirk van Greuning Founder, Growth Pulse Media

Founder of Growth Pulse Media and a specialist in South African search dominance. Dirk translates his experience in scaling South African businesses into high-velocity digital strategies for B2B and retail leaders. He writes about SEO, lead generation, and paid media from an operator's perspective — prioritising pipeline value over impressions.

Connect on LinkedIn