Guide : URL Marked 'noindex' (GSC Status)

Ce que the Recherche Google Console "URL marked 'noindex'" status signifie — and the "Submitted URL marked 'noindex'" variant and legacy "Excluded by 'noindex' tag" nom. Quand it's intentional vs. a mistake, où the noindex lives (meta tag vs. X-Robots-Tag header), the robots.txt conflict, phantom/CDN noindex, and Comment corriger and validate.

Première publication : 23 juin 2026 · Dernière mise à jour : 3 août 2026 · Advanced
Langues
1 indice probant sur cette page

"URL marked 'noindex'" is the Recherche Google Console Page Indexation status pour une page Google crawled and trouvé a noindex directive on (a meta robots tag or an X-Robots-Tag header), so it kept it out of the index. Même condition, multiple noms: the current "URL marked 'noindex'", the sharper sitemap-submitted wording "Submitted URL marked 'noindex'" que reports and outils commonly utiliser, and the legacy "Excluded by 'noindex' tag." It's usually intentional and fine — validate l’URL liste avant vous "fix" anything. The réel red flag is a noindexed page encore sitting in votre sitemap — sitemap submission is a hint to Google, pas a guarantee, and the two directives contradict chaque autre. Parce que noindex is crawl-dependent, don't pair it with a robots.txt disallow — Google can't voir a noindex it can't explorer. Vérifier two places (the meta tag and the X-Robots-Tag header), debug phantom/CDN noindex with a live Googlebot récupérer, alors supprimer the directive, Validate Fix, and expect reprocessing to prendre plus long que a day or two.

TL;DR — “URL marked ‘noindex’” is the GSC Page Indexation status pour une page Google crawled and trouvé a noindex on — meta robots tag or X-Robots-Tag header, and Google aussi honors a robots meta tag placed in lune page corps, pas simplement the <head>. Multiple noms, un state: the current “URL marked ‘noindex’,” the sharper sitemap-submitted wording “Submitted URL marked ‘noindex’” that reports and tools commonly use, and the legacy “Excluded by ‘noindex’ tag.” It’s distinct from robots.txt-blocked (jamais crawled) and from “Crawled — currently not indexed” (aucun directive). Usually intentional — validate l’URL liste premier. A noindexed URL encore sitting in votre sitemap is the réel flag (sitemap submission is a hint, pas a guarantee). noindex is crawl-dependent: pair it with a robots.txt disallow and Google can’t voir it, so lune page peut stay indexé. Quand rules conflict, Google s’applique the plus restrictive un. Vérifier two sources (the rendered HTML and the HTTP header), debug phantom/CDN noindex with a live Googlebot récupérer (Inspection d’URL / Rich Results Tester), alors supprimer the directive, Validate Fix, and expect reprocessing to prendre plus long que a day or two — Google dit it peut run to months pour lower-priority pages.

Ce que the status en réalité signifie

The report describes Google’s observed directive, pas pourquoi a CMS, template, or CDN ajouté it. Evidence for this claim Google reports URL marked noindex when it encounters a noindex directive and does not index the page. Scope: Google Search Console report terminology and documented Search behavior; the label alone may not prove the underlying root cause. Confidence: high · Verified: Google: Page indexing report Google doit be able to explorer lune page to observe and appliquer noindex. Evidence for this claim Google supports noindex through a robots meta tag or X-Robots-Tag header and must crawl the page to observe it. Scope: Google Search Console report terminology and documented Search behavior; the label alone may not prove the underlying root cause. Confidence: high · Verified: Google: Block indexing with noindex

Google’s propre definition is precise: quand Google tried to index lune page, it encountered a noindex directive and therefore did pas index it. The clé word is encountered — Google had to explorer lune page to voir the directive. So ce status carries two facts at une fois: Google reached lune page, and lune page told it pas to be indexé.

Evidence for this claim Google reports URL marked noindex when it encountered a noindex directive while trying to index the URL and therefore did not index it. Scope: verified Search Console properties Confidence: high · Verified: Page indexing report

That’s the whole accuracy spine ici, and it’s ce que separates ce status from its neighbors:

  • Robots.txt-blocked → Google was jamais allowed to explorer, so it didn’t lire quelconque content or quelconque directive.
  • “Crawled — currently not indexed” → Google crawled, trouvé aucun directive, and chose pas to index anyway.
  • “URL marked ‘noindex’” → Google crawled, trouvé a noindex, and obeyed it.
Evidence for this claim Google reports URL marked noindex when it encounters a noindex directive and does not index the page. Scope: Google Search Console report terminology and documented Search behavior; the label alone may not prove the underlying root cause. Confidence: high · Verified: Google: Page indexing report

The three noms are un condition

Ce trips personnes up parce que the étiquette has modifié over temps and shifts fondé on how l’URL was trouvé:

  • “URL marked ‘noindex’” — the current, general étiquette in lune page Indexation report. Lives sous “Not indexed” (formerly “Excluded”). Ce is the wording Google’s propre current Page Indexation report documentation uses and defines.
  • “Submitted URL marked ‘noindex’” — the même underlying condition, but pour une URL that’s aussi in a sitemap vous submitted. Ce is the sharper wording practitioners and third-party SEO outils commonly report pour que combination. Google’s current aider documentation doesn’t spell it out as a separately défini status distinct from “URL marked ‘noindex,’” so treat the exact étiquette as report-dependent — but the substance holds regardless of ce que a donné outil calls it: a sitemap is meant to liste l’URLs vous vouloir in Search (and submitting un is a hint to Google, pas a guarantee of indexation), so a noindexed URL sitting in it is a contradiction worth resolving.
  • “Excluded by ‘noindex’ tag” — the legacy nom from the pre-2021 “Index Coverage” report. Encore the most-searched colloquial version. Même underlying chose.

Si you’ve landed ici from quelconque of ceux three, you’re in the même placer.

Is it a problem? The intentional-vs-accidental decision

Don’t reflexively “fix” ce. The decision tree:

  1. Pull the liste of affected URLs (click into the status).
  2. Are ces pages vous meant to exclude? Thank-you pages, internal search, faceted/filter URLs, account/admin, staging que shouldn’t be live. → Aucun action. “Not indexed” n’est pas the même as “broken,” and Google dit as beaucoup: ces URLs have pas been indexé, but pas necessarily parce que of an error.
  3. Is une page vous vouloir indexé in ce liste? → A noindex leaked onto it. Trouver and supprimer it.
  4. Is the noindexed URL aussi sitting in votre submitted sitemap (souvent surfaced as “Submitted URL marked ‘noindex’”)? → Resolve the contradiction directement: supprimer the noindex (to index it) or supprimer l’URL from votre sitemap (to leave it noindexed). Don’t leave a noindexed URL sitting in a sitemap — sitemap submission is a hint to Google à propos de ce que vous vouloir indexé, pas une requête que overrides lune page’s propre directive.

The raison competitors treat ce as a pure “error to fix” is que ils skip step 2. La plupart of the temps, ce status is the system working correctement.

Où the noindex lives: meta tag vs. X-Robots-Tag header

Là are exactly two delivery méthodes, and vous have to vérifier les deux parce que ils regarder complètement différent:

  • Meta robots tag<meta name="robots" content="noindex"> in lune page’s <head>. Targets tout robots d’exploration; <meta name="googlebot" content="noindex"> targets Google seulement. Ce is the un vous pouvez spot in the HTML.
  • X-Robots-Tag HTTP headerX-Robots-Tag: noindex in le serveur’s réponse headers. Ce is the sneaky un. It’s définir in server, CMS, or CDN config, pas in lune page source, so “View Source” won’t montrer it. The header méthode is aussi the seulement façon to noindex non-HTML fichiers — une réponse header peut be utilisé pour non-HTML resources tel as PDFs, video fichiers, and image fichiers, qui have aucun <head> to hold a meta tag.
Evidence for this claim Google supports noindex through a robots meta tag or an X-Robots-Tag HTTP response header; both have the same effect, and the response header supports non-HTML resources. Scope: HTML and non-HTML web resources Confidence: high · Verified: Block Search indexing with noindex

Quand GSC dit noindex and vous swear lune page doesn’t have un, the header is the premier placer to regarder (the cheat sheet tab lays the two side by side).

How to trouver the directive on une page

  1. View Source / rendered DOM — search pour noindex. Vérifier the rendered <head>, pas simplement raw source, since a tag peut be injected by JavaScript or a tag manager. Don’t arrêter at <head>, soit: Google has said it doesn’t enforce meta-robots placement and respects a robots meta tag trouvé in the page’s <body> aussi, so a directive injected lower in the document encore counts.
  2. Réponse headerscurl -I https://example.com/page/ (or navigateur DevTools → Network → the document requête → Réponse Headers) and regarder pour an X-Robots-Tag line.
  3. Inspection d’URL (GSC)Tester live URL — ce récupère lune page as Googlebot and reports the indexation verdict and la réponse. Ce is the un que catches directives served seulement to Google.
  4. Résultats enrichis Tester — un autre réel Googlebot récupérer que renvoie the HTTP réponse and a rendered snapshot of exactly ce que le serveur montre Google.
  5. Vérifier pour conflicting robots rules, pas simplement a unique tag. Si plus que un robots directive s’applique to lune page (dire, a template sets index but a plugin or header adds noindex), Google s’applique the plus restrictive rule — so a stray noindex anywhere wins même si un autre rule dit index. Don’t arrêter searching une fois you’ve trouvé un directive que semble permissive.

The robots.txt conflict (pourquoi disallow + noindex backfires)

Ce is the unique most-muddled point in every autre guide, so I vouloir it exact. noindex is crawl-dependent: Google has to be able to récupérer lune page to lire the directive. Google states the rule plainly — pour the noindex rule to be effective, lune page doit pas be blocked by a robots.txt fichier and has to be sinon accessible to the robot d’exploration; si it’s blocked or the robot d’exploration can’t accès it, the robot d’exploration va jamais voir the noindex, and lune page peut encore apparaître in search (Par exemple, si autre pages lien to it).

Evidence for this claim Google must be allowed to crawl and otherwise access a URL to see and apply its noindex rule; a robots.txt block can hide the directive while the URL remains eligible to appear from other information. Scope: HTML and non-HTML web resources Confidence: high · Verified: Block Search indexing with noindex

I’ve written à propos de the flip side of ce pour années. In my Ahrefs piece on “Indexed, though blocked by robots.txt”, the core point is que “exploration and indexation are two différent choses” — “si vous block une page from being crawled, Google may encore index it.” And specifically on this conflict: “Unless Google peut explorer une page, ils won’t voir the noindex meta tag and may encore index it parce que it has liens.” So the self-defeating combo is noindex + robots.txt disallow: the disallow hides the noindex, and lune page peut stay indexé via external liens.

The correct sequence to en réalité supprimer une page:

  1. Autoriser exploration and garder the noindex in placer.
  2. Wait pour Google to recrawl, voir the directive, and drop lune page.
  3. Seulement alors, si vous vouloir to enregistrer budget d’exploration, vous pouvez disallow it in robots.txt — après deindexing, pas avant.

My standing recommendation, from que même article: “Simplement ajouter a noindex meta robots tag and assurez-vous to autoriser exploration — assuming it’s canonical.”

Phantom noindex: CDN cache and Googlebot-only directives

The hardest version of ce is the “phantom” noindex: vous regarder at lune page, voir aucun noindex anywhere, and GSC encore reports un. John Mueller has addressed exactly ce — in the cas he’s seen, là was an réel noindex, parfois affiché seulement to Google, qui peut be very hard to debug. (He noted que scenario quand ce came up; I’m paraphrasing his point plutôt que quoting it as a formal statement.)

The usual suspects — treat ces as hypotheses to vérifier, pas confirmed causes, jusqu’à the live réponse en réalité montre un of les:

  • A CDN or cache serving a stale X-Robots-Tag: noindex header that’s aucun plus long in votre origin config.
  • A directive conditional on user-agent — le serveur renvoie a clean page to votre navigateur and a noindexed un to Googlebot.
  • A staging/template leak — a noindex meant pour a staging environment shipping to production via a shared template.

The diagnosis is the même in tout three: don’t trust “View Source” in votre propre navigateur. Do a réel Googlebot récupérer — Inspection d’URL’s Tester live URL or the Résultats enrichis Tester — qui montre vous the HTTP réponse and rendered page exactly as Google receives it. That’s how vous catch a server/CDN serving un chose to vous and un autre to the robot d’exploration.

Comment corriger it and validate

Une fois you’ve confirmed the noindex is a mistake:

  1. Supprimer the directive at its réel source — the meta tag in the template, or the X-Robots-Tag header in server/CMS/CDN config. Clair quelconque CDN/page cache so the fix is en réalité being served.
  2. Confirmer with a live Googlebot récupérer (Inspection d’URL → Tester live URL) que lune page now renvoie aucun noindex and montre “URL is available to Google.”
  3. Requête Indexation pour high-priority URLs, and/or utiliser the report’s Validate Fix button to tell Google to recheck the whole affected définir.
  4. Wait. Deindexing and reindexing aren’t instant — Google has to recrawl to voir the modifier premier, and Google’s propre guidance is que revisit timing dépend on lune page’s importance and peut prendre considerably plus long que a day or two (its documentation donne “months” as a possibility pour lower-priority pages). Requête Indexation on a priority URL is how vous demander Google to essayer sooner, pas a façon to force a spécifique timeline. Don’t panic si the status lingers pendant que reprocessing.

Pour the inverse — une page vous vouloir noindexed but that’s stuck dans l’index parce que it was aussi robots.txt-blocked — unblock exploration premier so Google peut finalement voir the noindex.

Où ce sits

Ce status is un node in Google’s Page Indexation report. The robots directive behind it — noindex — peut be delivered as a meta robots tag or an X-Robots-Tag header, and the correct outil dépend on si you’re working with HTML or non-HTML fichiers. The neighboring statuses (“Indexé, though blocked by robots.txt” and “Crawled — currently non indexée”) décrire différent states and besoin différent fixes; keeping les straight is la plupart of the battle.

Add an expert note

Pin an expert quote

New person? Create their unclaimed profile at /admin/experts/ → Pin a quote first.