URL mit Markierung 'noindex' (GSC-Status)
war the Google Search Console "URL marked 'noindex'" status bedeutet — und the "Submitted URL marked 'noindex'" variant und legacy "Excluded durch 'noindex' tag" name. wenn es ist intentional vs. ein mistake, wo the noindex lives (meta tag vs. X-Robots-Tag header), the robots.txt conflict, phantom/CDN noindex, und wie zu fix und validate.
Sprachen
1 Evidenzsignal auf dieser Seite
- Verknüpftes Live-WerkzeugGoogle Index Checker
"URL marked 'noindex'" ist the Google Search Console Seite Indexing status für ein Seite Google crawled und gefunden ein noindex directive auf (ein meta robots tag oder ein X-Robots-Tag header), so es kept es out von the index. gleich condition, multiple names: the current "URL marked 'noindex'", the sharper sitemap-submitted wording "Submitted URL marked 'noindex'" that Berichte und Tools commonly verwenden, und the legacy "Excluded durch 'noindex' tag." es ist usually intentional und fine — validate the URL Liste vor Sie "fix" anything. The real red flag ist ein noindexed Seite still sitting in Ihre sitemap — sitemap submission ist ein hint zu Google, nicht ein guarantee, und the two directives contradict jede other. weil noindex ist crawlen-dependent, don't pair es mit ein robots.txt disallow — Google kann nicht sehen ein noindex es kann nicht crawlen. prüfen two places (the meta tag und the X-Robots-Tag header), debug phantom/CDN noindex mit ein live Googlebot fetch, then entfernen the directive, Validate Fix, und expect reprocessing zu nehmen longer than ein day oder two.
Evidence for this claim Google reports URL marked noindex when it encounters a noindex directive and does not index the page. Scope: Google Search Console report terminology and documented Search behavior; the label alone may not prove the underlying root cause. Confidence: high · Verified: Google: Page indexing reportTL;DR — “URL marked ‘noindex’” in Google Search Console bedeutet Google looked bei Ihre Seite, saw ein
noindexinstruction auf es, und kept es out von Suche auf purpose. meisten von the time das ist intended — lots von Seiten sollte sein noindexed. The situation zu tatsächlich worry über ist ein Seite Sie put in Ihre sitemap (asking Google zu index es) that auch says don’t-index — einige Berichte call dies “Submitted URL marked ‘noindex’”. Look bei the Liste von affected Seiten: wenn sie sind alle ones Sie meant zu hide, Sie sind done.
war dies status ist telling Sie
dies label bedeutet Google encountered ein noindex directive während processing the Seite. Evidence for this claim Google reports URL marked noindex when it encounters a noindex directive and does not index the page. Scope: Google Search Console report terminology and documented Search behavior; the label alone may not prove the underlying root cause. Confidence: high · Verified: Google: Page indexing report Google supports noindex in ein robots meta element oder X-Robots-Tag response header. Evidence for this claim Google supports noindex through a robots meta tag or X-Robots-Tag header and must crawl the page to observe it. Scope: Google Search Console report terminology and documented Search behavior; the label alone may not prove the underlying root cause. Confidence: high · Verified: Google: Block indexing with noindex
wenn Sie offen the Seite Indexing Bericht in Search Console und sehen ein row called
“URL marked ‘noindex’”, hier’s war happened: Google visited (crawled) the
Seite, gefunden ein noindex instruction auf es, und decided nicht zu put es in Suche
Ergebnisse. das ist es. The Seite ist nicht broken — Google tat exactly war the Seite
told es zu tun.
ein noindex ist ein small instruction that says “don’t zeigen dies Seite in Suche.” es
lives in one von two places:
- ein line in the Seite’s code (ein meta robots tag), oder
- ein setting in the Seite’s server response (ein X-Robots-Tag header) — dies one Sie kann nicht sehen durch just looking bei the Seite.
ist dies bad? Usually nicht
The word that scares Menschen ist “nicht indexed” (the section dies status lives under). es sounds like ein error. es usually ist nicht. Plenty von Seiten sollte sein kept out von Suche:
- Thank-Sie / order-confirmation Seiten
- intern Suchergebnisse
- Login, account, und admin Seiten
- Filtered oder sorted versions von ein listing
wenn the Seiten in dies Liste sind ones Sie meant zu hide, leave them alone. es gibt nothing zu fix.
wenn Sie tun benötigen zu act
Two situations:
- ein Seite Sie wollen in Google ist in dies Liste. Something put ein
noindexauf ein Seite that sollte ranken. das ist ein mistake zu verfolgen down. - Sie sehen “Submitted URL marked ‘noindex’” (ein slightly different, sharper
wording einige Berichte und Tools verwenden). Either Weg, the substance ist the gleich:
Sie submitted the Seite in Ihre sitemap — welche ist meant zu Liste Seiten Sie
wollen in Suche — aber the Seite auch says “don’t index.” diese two things
contradict jede other. Either entfernen the
noindex(wenn Sie wollen es indexed) oder nehmen the URL out von Ihre sitemap (wenn Sie don’t).
Die Namensverwechslung
Sie might auch remember dies als “Excluded durch ‘noindex’ tag.” das ist just the older name für the gleich thing aus Google’s previous Bericht. So three labels — “URL marked ‘noindex’,” “Submitted URL marked ‘noindex’,” und “Excluded durch ‘noindex’ tag” — alle describe one situation: Google gefunden ein noindex.
One trap zu know über
ein common instinct ist zu auch block the Seite in robots.txt zu “really” halten es
out. Don’t. Blocking crawling stops Google aus reading the Seite bei alle — welche
bedeutet es kann nicht sehen Ihre noindex either. Counterintuitively, the Seite kann stay
in Suche. zu entfernen ein Seite, let Google crawlen es und halten the noindex auf es.
wollen the full diagnostic version — meta tag vs. header detection, phantom/CDN noindex, und the fix-und-validate flow — switch zu the Fortgeschritten tab.
TL;DR — “URL marked ‘noindex’” ist the GSC Seite Indexing status für ein Seite Google crawled und gefunden ein
noindexauf — meta robots tag oder X-Robots-Tag header, und Google auch honors ein robots meta tag placed in the Seite body, nicht just the<head>. Multiple names, one state: the current “URL marked ‘noindex’,” the sharper sitemap-submitted wording “Submitted URL marked ‘noindex’” that Berichte und Tools commonly verwenden, und the legacy “Excluded durch ‘noindex’ tag.” es ist distinct aus robots.txt-blocked (never crawled) und aus “Crawled — currently nicht indexed” (kein directive). Usually intentional — validate the URL Liste erste. ein noindexed URL still sitting in Ihre sitemap ist the real flag (sitemap submission ist ein hint, nicht ein guarantee).noindexist crawlen-dependent: pair es mit ein robots.txt disallow und Google kann nicht sehen es, so the Seite kann stay indexed. wenn rules conflict, Google applies the mehr restrictive one. prüfen two Quellen (the rendered HTML und the HTTP header), debug phantom/CDN noindex mit ein live Googlebot fetch (URL Inspection / Rich Ergebnisse testen), then entfernen the directive, Validate Fix, und expect reprocessing zu nehmen longer than ein day oder two — Google says es kann ausführen zu months für lower-priority Seiten.
war the status tatsächlich bedeutet
The Bericht describes Google’s observed directive, nicht warum ein CMS, template, oder CDN hinzugefügt es. Evidence for this claim Google reports URL marked noindex when it encounters a noindex directive and does not index the page. Scope: Google Search Console report terminology and documented Search behavior; the label alone may not prove the underlying root cause. Confidence: high · Verified: Google: Page indexing report Google must sein able zu crawlen the Seite zu observe und anwenden noindex. Evidence for this claim Google supports noindex through a robots meta tag or X-Robots-Tag header and must crawl the page to observe it. Scope: Google Search Console report terminology and documented Search behavior; the label alone may not prove the underlying root cause. Confidence: high · Verified: Google: Block indexing with noindex
Google’s own definition ist precise: wenn Google tried zu index the Seite, es
encountered ein noindex directive und therefore tat nicht index es. The key word ist
encountered — Google hatte zu crawlen the Seite zu sehen the directive. So dies
status carries two facts bei once: Google reached the Seite, und the Seite told es
nicht zu sein indexed.
das ist the whole accuracy spine hier, und es ist war separates dies status aus its neighbors:
- Robots.txt-blocked → Google war never allowed zu crawlen, so es didn’t lesen any Inhalt oder any directive.
- “Crawled — currently nicht indexed” → Google crawled, gefunden kein directive, und chose nicht zu index anyway.
- “URL marked ‘noindex’” → Google crawled, gefunden ein
noindex, und obeyed es.
The three names sind one condition
dies trips Menschen up weil the label hat changed over time und shifts auf Grundlage von wie the URL war gefunden:
- “URL marked ‘noindex’” — the current, general label in the Seite Indexing Bericht. Lives under “nicht indexed” (formerly “Excluded”). dies ist the wording Google’s own current Page Indexing report documentation uses und defines.
- “Submitted URL marked ‘noindex’” — the gleich underlying condition, aber für ein URL das ist auch in ein sitemap Sie submitted. dies ist the sharper wording practitioners und Drittanbieter- SEO Tools commonly Bericht für that combination. Google’s current helfen documentation tut nicht spell es out als ein separately defined status distinct aus “URL marked ‘noindex,’” so treat the exact label als Bericht-dependent — aber the substance holds regardless von war ein given Tool calls es: ein sitemap ist meant zu Liste the URLs Sie wollen in Suche (und submitting one ist ein hint zu Google, nicht ein guarantee von indexing), so ein noindexed URL sitting in es ist ein contradiction worth resolving.
- “Excluded durch ‘noindex’ tag” — the legacy name aus the pre-2021 “Index Coverage” Bericht. Still the meisten-searched colloquial version. gleich underlying thing.
wenn Sie’ve landed hier aus any von diese three, Sie sind in the gleich place.
ist es ein problem? The intentional-vs-accidental decision
Don’t reflexively “fix” dies. The decision tree:
- Pull the Liste von affected URLs (klicken into the status).
- sind these Seiten Sie meant zu exclude? Thank-Sie Seiten, intern Suche, faceted/filter URLs, account/admin, staging that shouldn’t sein live. → kein action. “nicht indexed” ist nicht the gleich als “broken,” und Google says als much: these URLs haben nicht been indexed, aber nicht necessarily weil von ein error.
- ist ein Seite Sie wollen indexed in dies Liste? → ein noindex leaked onto es. finden und entfernen es.
- ist the noindexed URL auch sitting in Ihre submitted sitemap (häufig
surfaced als “Submitted URL marked ‘noindex’”)? → lösen the contradiction
directly: entfernen the
noindex(zu index es) oder entfernen the URL aus Ihre sitemap (zu leave es noindexed). Don’t leave ein noindexed URL sitting in ein sitemap — sitemap submission ist ein hint zu Google über war Sie wollen indexed, nicht ein Anfrage that overrides the Seite’s own directive.
The Grund competitors treat dies als ein pure “error zu fix” ist that they skip step 2. meisten von the time, dies status ist the System working correctly.
wo the noindex lives: meta tag vs. X-Robots-Tag header
dort sind exactly two delivery methods, und Sie haben zu prüfen both weil they look completely different:
- Meta robots tag —
<meta name="robots" content="noindex">in the Seite’s<head>. Targets alle crawlers;<meta name="googlebot" content="noindex">targets Google nur. dies ist the one Sie kann spot in the HTML. - X-Robots-Tag HTTP header —
X-Robots-Tag: noindexin the server’s response headers. dies ist the sneaky one. es ist festlegen in server, CMS, oder CDN config, nicht in the Seite Quelle, so “View Quelle” wird nicht zeigen es. The header Methode ist auch the nur Weg zu noindex non-HTML files — ein response header kann sein verwendet für non-HTML resources such als PDFs, video files, und image files, welche haben kein<head>zu hold ein meta tag.
wenn GSC says noindex und Sie swear the Seite tut nicht haben one, the header ist the erste place zu look (the cheat sheet tab lays the two side durch side).
wie zu finden the directive auf ein Seite
- View Quelle / rendered DOM — Suche für
noindex. prüfen the rendered<head>, nicht just raw Quelle, since ein tag kann sein injected durch JavaScript oder ein tag manager. Don’t stop bei<head>, either: Google hat said es tut nicht enforce meta-robots placement und respects ein robots meta tag gefunden in the Seite’s<body>too, so ein directive injected lower in the Dokument still counts. - Response headers —
curl -I https://example.com/page/(oder browser DevTools → Network → the Dokument Anfrage → Response Headers) und look für einX-Robots-Tagline. - URL Inspection (GSC) → testen live URL — dies fetches the Seite als Googlebot und Berichte the indexing verdict und the response. dies ist the one that catches directives served nur zu Google.
- Rich Results testen — another real Googlebot fetch that returns the HTTP response und ein rendered snapshot von exactly war the server zeigt Google.
- prüfen für conflicting robots rules, nicht just ein single tag. wenn mehr als
one robots directive applies zu the Seite (say, ein template sets
indexaber ein plugin oder header addsnoindex), Google applies the mehr restrictive rule — so ein straynoindexanywhere wins even wenn another rule saysindex. Don’t stop searching once Sie’ve gefunden one directive that looks permissive.
The robots.txt conflict (warum disallow + noindex backfires)
dies ist the single meisten-muddled point in every other Leitfaden, so I wollen es exact.
noindex ist crawlen-dependent: Google hat zu sein able zu fetch the Seite zu lesen
the directive. Google states the rule plainly — für the noindex rule zu sein
effective, the Seite must nicht sein blocked durch ein robots.txt file und hat zu sein
otherwise accessible zu the crawler; wenn es ist blocked oder the crawler kann nicht access
es, the crawler will never sehen the noindex, und the Seite kann still erscheinen in
Suche (zum Beispiel, wenn other Seiten Link zu es).
I’ve written über the flip side von dies für Jahre. in my Ahrefs piece auf “Indexed, though blocked by robots.txt”, the core point ist that “crawling und indexing sind two different things” — “wenn Sie block ein Seite aus being crawled, Google may still index es.” und specifically auf dies conflict: “Unless Google kann crawlen ein Seite, they wird nicht sehen the noindex meta tag und may still index es weil es hat Links.” So the self-defeating combo ist noindex + robots.txt disallow: the disallow hides the noindex, und the Seite kann stay indexed via external Links.
The correct sequence zu tatsächlich entfernen ein Seite:
- erlauben crawling und halten the
noindexin place. - Wait für Google zu recrawl, sehen the directive, und drop the Seite.
- nur then, wenn Sie wollen zu speichern crawl budget, Sie kann disallow es in robots.txt — after deindexing, nicht vor.
My standing Empfehlung, aus that same article: “Just hinzufügen ein noindex meta robots tag und machen sure zu erlauben crawling — assuming es ist canonical.”
Phantom noindex: CDN cache und Googlebot-nur directives
The hardest version von dies ist the “phantom” noindex: Sie look bei the Seite, sehen kein noindex anywhere, und GSC still Berichte one. John Mueller hat addressed exactly dies — in the cases he’s seen, dort war ein actual noindex, sometimes shown nur zu Google, welche kann sein very hard zu debug. (He noted that scenario wenn dies came up; I’m paraphrasing his point anstatt quoting es als ein formal statement.)
The usual suspects — treat these als hypotheses zu prüfen, nicht confirmed causes, until the live response tatsächlich zeigt one von them:
- ein CDN oder cache serving ein stale
X-Robots-Tag: noindexheader das ist kein longer in Ihre origin config. - ein directive conditional auf user-agent — the server returns ein clean Seite zu Ihre browser und ein noindexed one zu Googlebot.
- ein staging/template leak — ein noindex meant für ein staging environment shipping zu production durch ein shared template.
The diagnosis ist the gleich in alle three: don’t trust “View Quelle” in Ihre own browser. tun ein real Googlebot fetch — URL Inspection’s testen live URL oder the Rich Results testen — welche zeigt Sie the HTTP response und rendered Seite exactly als Google receives es. das ist wie Sie catch ein server/CDN serving one thing zu Sie und another zu the crawler.
wie zu fix es und validate
Once Sie’ve confirmed the noindex ist ein mistake:
- entfernen the directive bei its real Quelle — the meta tag in the template, oder
the
X-Robots-Tagheader in server/CMS/CDN config. klar any CDN/page cache so the fix ist tatsächlich being served. - Confirm mit ein live Googlebot fetch (URL Inspection → testen live URL) that the Seite now returns kein noindex und zeigt “URL ist verfügbar zu Google.”
- Anfrage Indexing für high-priority URLs, and/or verwenden the Bericht’s Validate Fix button zu mitteilen Google zu recheck the whole affected festlegen.
- Wait. Deindexing und reindexing sind nicht instant — Google hat zu recrawl zu sehen the ändern erste, und Google’s own guidance ist that revisit timing depends auf the Seite’s importance und kann nehmen considerably longer than ein day oder two (its documentation gives “months” als ein possibility für lower-priority Seiten). Anfrage Indexing auf ein priority URL ist wie Sie fragen Google zu try sooner, nicht ein Weg zu force ein specific timeline. Don’t panic wenn the status lingers während reprocessing.
für the inverse — ein Seite Sie wollen noindexed aber das ist stuck in the index
weil es war auch robots.txt-blocked — unblock crawling erste so Google kann
finally sehen the noindex.
wo dies sits
dies status ist one node in Google’s Seite Indexing Bericht. The robots
directive behind es — noindex — kann sein delivered als ein meta robots tag oder ein
X-Robots-Tag header, und the right Tool depends auf whether Sie sind working mit
HTML oder non-HTML files. The neighboring statuses (“Indexed, though blocked durch
robots.txt” und “Crawled — currently nicht indexed”) describe different states und
benötigen different fixes; keeping them straight ist meisten von the battle.
AI summary
ein condensed nehmen auf the Advanced version:
- war es bedeutet: “URL marked ‘noindex’” ist the GSC Seite Indexing status für ein
Seite Google crawled und gefunden ein
noindexauf — so es kept es out von the index auf purpose. Google hatte zu crawlen the Seite zu sehen the directive, und es honors that directive whether es ist in the<head>oder the Seite body. - Multiple names, one state: the current “URL marked ‘noindex’,” the sharper sitemap-submitted wording “Submitted URL marked ‘noindex’” that Berichte und Tools commonly verwenden, und the legacy “Excluded durch ‘noindex’ tag.”
- Distinct aus neighbors: robots.txt-blocked = never crawled; “Crawled — currently nicht indexed” = crawled, kein directive; dies status = crawled, gefunden ein noindex, obeyed es.
- Usually intentional. Validate the URL Liste erste. Thank-Sie Seiten, intern Suche, facets, admin = fine, kein action. nur act wenn ein Seite Sie wollen indexed ist in the Liste — oder the URL ist noindexed aber still sitting in Ihre sitemap (ein contradiction, since sitemap submission ist ein hint für war Sie wollen indexed, nicht ein guarantee): entfernen the noindex oder entfernen es aus the sitemap.
- Two delivery methods: ein meta robots tag anywhere in the rendered HTML
(nicht just the
<head>), oder ein X-Robots-Tag HTTP header (the nur Weg für non-HTML files like PDFs, und the sneaky one — nicht in Seite Quelle). prüfen both, und remember that wenn multiple robots rules conflict, Google applies the mehr restrictive one. - The robots.txt conflict:
noindexist crawlen-dependent. Disallow + noindex backfires — Google kann nicht sehen ein noindex es kann nicht crawlen, so the Seite kann stay indexed via Links. zu entfernen ein Seite: erlauben crawling + halten noindex, then optionally disallow after deindexing. - Phantom noindex: owner sees none, Google tut — usually ein CDN/cache serving ein stale header, ein Googlebot-nur directive, oder ein staging/template leak (treat these als hypotheses zu confirm, nicht assumed causes). Debug mit ein real Googlebot fetch (URL Inspection live testen / Rich Results testen), nicht View Quelle.
- Fix → validate: entfernen the directive bei its Quelle, klar caches, confirm via ein live Googlebot fetch, Anfrage Indexing / Validate Fix, then wait — Google’s own guidance says reprocessing depends auf the Seite’s importance und kann nehmen much longer than ein day oder two, bis zu months für lower-priority Seiten.
Offizielle Dokumentation
Primary-Quelle documentation aus the Suchmaschinen.
- Page Indexing report — the Bericht dies status lives in; defines “URL marked ‘noindex’” und the related “Indexed, though blocked by robots.txt” (its current text tut nicht separately spell out “Submitted URL marked ‘noindex’” als ein distinct status name, though the underlying sitemap contradiction es describes ist real).
- Block search indexing with noindex — war
noindextut, the meta-tag vs. X-Robots-Tag header methods, the crawlen-dependent rule (ein blocked Seite never sees the noindex), und Google’s own note that revisiting ein Seite after ein ändern kann nehmen months depending auf its importance. - Robots meta tag, data-nosnippet, and X-Robots-Tag specifications — wie conflicting robots rules lösen (the mehr restrictive rule wins) und confirmation that Google auch respects ein robots meta tag placed in the Seite body, nicht just the
<head>. - Introduction to robots.txt — warum robots.txt controls crawling, nicht indexing, und warum es ist nicht ein deindexing Tool.
- URL Inspection tool — “testen live URL” fetches the Seite als Googlebot, the Weg zu catch directives served nur zu Google.
- Build and submit a sitemap — sitemaps sollte Liste the URLs Sie wollen in Suche, und submitting one ist ein hint zu Google, nicht ein guarantee von crawling oder indexing.
Bing / Microsoft
- Bing Webmaster Tools — Help & How-To — Bing honors the robots
<meta name="robots" content="noindex">tag und theX-Robots-Tagheader the gleich Weg; its index Berichte surface noindexed Seiten similarly. (Lower priority für dies Google-specific status.)
Quotes aus the Quelle
auf-the-record statements aus Google, plus my own writing auf the robots.txt conflict. jede Link ist ein deep Link that jumps zu the quoted passage auf the Quelle Seite.
Google — war the status bedeutet
- “wenn Google tried zu index the Seite es encountered ein ‘noindex’ directive und therefore tat nicht index es.” — Google Search Console helfen (Page Indexing report). Jump to quote
Google — noindex ist crawlen-dependent
- “für the noindex rule zu sein effective, the Seite oder resource must nicht sein blocked durch ein robots.txt file, und es hat zu sein otherwise accessible zu the crawler. wenn the Seite ist blocked durch ein robots.txt file oder the crawler kann nicht access the Seite, the crawler will never sehen the noindex rule, und the Seite kann still erscheinen in Suchergebnisse, zum Beispiel wenn other Seiten Link zu es.” — Google Suche Central docs (Block Suche indexing mit noindex). Jump to quote
- “Depending auf the importance von the Seite auf the internet, es may nehmen months für Googlebot zu revisit ein Seite.” — Google Suche Central docs (Block Suche indexing mit noindex), auf wie long reprocessing after ein noindex ändern kann nehmen. Read the article
Google — conflicting rules und wo ein directive kann live
- “Google Suche tut nicht enforce placement von meta robots in the HTML head und will respect robots meta tags in the body section von ein HTML Dokument als well.” — Google Suche Central docs (Robots meta tag, Daten-nosnippet, und X-Robots-Tag specifications). Read the article
- “in the case von conflicting robots rules, the mehr restrictive rule applies.” — Google Suche Central docs (gleich specification). Read the article
Patrick Stox (Ahrefs) — der Konflikt mit robots.txt
- “wenn Sie block ein Seite aus being crawled, Google may still index es weil crawling und indexing sind two different things.” — me, “Indexed, though blocked by robots.txt” kann sein mehr als ein Robots.txt Block (Ahrefs). Jump to quote
- “Unless Google kann crawlen ein Seite, they wird nicht sehen the noindex meta tag und may still index es weil es hat Links.” — me (gleich article). Read the article
- “Just hinzufügen ein noindex meta robots tag und machen sure zu erlauben crawling—assuming es ist canonical.” — me (gleich article). Read the article
Checkliste für „URL mit Markierung ‘noindex’“
arbeiten top zu bottom — meisten von these end bei step 2 mit “kein action needed.”
- offen the affected URL Liste in the Page Indexing report (klicken the status row).
- entscheiden intentional vs. accidental: sind these Seiten Sie meant zu exclude (thank-Sie, intern Suche, facets, admin, staging)? wenn yes → done.
- ist the URL auch sitting in Ihre submitted sitemap (häufig surfaced als
“Submitted URL marked ‘noindex’”)? lösen the contradiction: entfernen the
noindex(zu index) oder entfernen the URL aus the sitemap (zu halten es out). - für Seiten that sollte sein indexed, finden the directive in both places:
- Rendered
<head>für<meta name="robots" ... noindex>(prüfen the rendered DOM, nicht just View Quelle). - Response headers für
X-Robots-Tag: noindex(curl -Ioder DevTools → Network).
- Rendered
- Confirm war Googlebot sees mit URL Inspection → testen live URL (oder the Rich Results testen) — catches Googlebot-nur / CDN-served directives.
- prüfen für the robots.txt conflict: the URL ist nicht auch disallowed (ein disallow hides the noindex und kann leave the Seite indexed).
- entfernen the directive bei its real Quelle (template / server / CMS / CDN), then klar any CDN oder Seite cache.
- Re-testen live that the Seite now returns kein noindex.
- Anfrage Indexing and/or hit Validate Fix; then wait — recrawl timing ist nicht fixed. Google says es depends auf the Seite’s importance und kann nehmen much longer than ein day oder two, bis zu months für lower-priority Seiten.
Cheat sheets
The three names — gleich condition
| Label Sie saw | wenn es zeigt | Severity |
|---|---|---|
| URL marked ‘noindex’ | Google crawled the Seite und gefunden ein noindex | Info (under “nicht indexed”) — häufig intentional |
| Submitted URL marked ‘noindex’ | gleich, aber the URL war in ein submitted sitemap | häufig surfaced als ein error-level row — regardless von label, ein contradiction zu lösen |
| Excluded durch ‘noindex’ tag | Legacy name (pre-2021 Index Coverage report) | gleich als “URL marked ‘noindex’” |
wo the noindex lives: meta tag vs. X-Robots-Tag header
| Meta robots tag | X-Robots-Tag header | |
|---|---|---|
| Form | <meta name="robots" content="noindex"> | X-Robots-Tag: noindex |
| Lives in | The Seite’s <head> (HTML) | The HTTP response headers |
| Visible in View Quelle? | Yes (wenn nicht JS-injected) | kein — prüfen curl -I / DevTools |
| festlegen durch | Template / CMS / Seite editor | Server / CMS / CDN config |
| funktioniert für non-HTML (PDF, image, video)? | kein (kein <head>) | Yes |
| Target one Engine? | name="googlebot" etc. | X-Robots-Tag: googlebot: noindex |
wie dies status differs aus its neighbors
| Status | Crawled? | Directive gefunden? | Meaning |
|---|---|---|---|
| URL marked ‘noindex’ | Yes | noindex | Google obeyed Ihre noindex |
| Indexed, though blocked by robots.txt | kein | n/a (kann nicht lesen es) | Blocked aus crawlen aber indexed via Links |
| Crawled — currently nicht indexed | Yes | None | kein directive; Google just chose nicht zu index |
Entscheidungsbaum: beabsichtigt vs. versehentlich
| Question | wenn yes | wenn kein |
|---|---|---|
| sind these Seiten Sie meant zu exclude? | kein action | go down ↓ |
| ist es the “Submitted” error variant? | entfernen noindex oder entfernen aus sitemap | go down ↓ |
| tun Sie wollen dies Seite indexed? | finden + entfernen the noindex, then Validate Fix | Leave es (und entfernen aus sitemap wenn present) |
Die mentalen Modelle
1. Crawled-und-saw-es. dies status nur exists weil Google crawled the Seite und lesen ein directive. That single fact distinguishes es aus robots.txt-blocked (never crawled) und aus “Crawled — currently nicht indexed” (crawled, kein directive). Locate welche von the three Sie sind in vor Sie touch anything.
2. “nicht indexed” ≠ broken. The Standard assumption sollte sein intentional, nicht error. Validate the Liste von URLs erste; meisten von the time the right move ist zu tun nothing. The one combination worth treating als urgent ist ein noindexed URL das ist auch sitting in Ihre submitted sitemap — weil ein sitemap ist meant zu Liste war Sie wollen indexed (submission ist nur ein hint zu Google, nicht ein guarantee), und ein noindexed sitemap URL contradicts that.
3. Two Quellen, immer prüfen both.
ein noindex ist either ein meta tag (in the rendered <head>) oder ein X-Robots-Tag
header (in the HTTP response). The header ist invisible in View Quelle, so “I
don’t haben ein noindex” usually bedeutet “I didn’t prüfen the header.” prüfen both, every
time.
4. crawlen-dependent — never pair noindex mit ein disallow. Google must crawlen ein Seite zu sehen its noindex. Block crawling in robots.txt und the noindex wird invisible, leaving the Seite indexable via Links. zu entfernen ein Seite: erlauben crawling + halten noindex, wait für deindexing, then optionally disallow.
5. Trust Googlebot’s view, nicht Ihre browser’s. für phantom noindex (Sie sehen none, GSC sees one), Ihre browser’s View Quelle ist the wrong instrument. ein CDN, cache, oder user-agent rule kann zeigen Googlebot something different. Diagnose mit ein real Googlebot fetch — URL Inspection’s live testen oder the Rich Results testen — welche zeigt the exact response Google receives.
Playbook: Sie just saw ein “marked ‘noindex’” status
lesen dies top zu bottom the erste time Sie hit one von these statuses. Branch off bei jede “wenn Sie sehen” — meisten läuft end early.
1. offen the Page Indexing report und klicken into the status row. Note welche von the three labels es ist: “URL marked ‘noindex’,” “Submitted URL marked ‘noindex,’” oder the legacy “Excluded durch ‘noindex’ tag.” Pull the full Liste von affected URLs — don’t judge aus the count alone.
2. Skim the URL Liste für shape. wenn Sie sehen mostly thank-Sie Seiten, intern Suche, filter/facet URLs, oder admin/account Seiten — these sind Seiten Sie’d normally wollen out von the index. → Stop hier. kein action needed; dies ist the System working als intended.
3. wenn Sie sehen ein URL Sie tatsächlich wollen Ranking, isolate es.
Someone oder something put ein noindex auf ein Seite that sollte sein indexable. Move zu
step 4 zu finden wo es ist coming aus.
4. wenn the label ist “Submitted URL marked ‘noindex’” — oder Sie just notice the
URL ist both noindexed und in Ihre sitemap — treat es als urgent.
Sitemap submission ist meant zu signal the URLs Sie wollen in Suche (Google
treats es als ein hint, nicht ein guarantee), so ein noindex auf that gleich URL ist ein
direct contradiction, und viele Berichte und Tools surface es als ein error-level
row für exactly that Grund. entscheiden welche side ist correct — Sie wollen es
indexed (entfernen the noindex) oder Sie don’t (pull es aus the sitemap) — und
lösen the contradiction the gleich day Sie finden es.
5. Locate the directive.
prüfen the rendered <head> für ein <meta name="robots" content="noindex"> tag,
then prüfen the response headers (curl -I oder DevTools → Network) für ein
X-Robots-Tag: noindex. wenn neither zeigt anything und GSC still Berichte one,
Sie likely haben ein phantom noindex — go zu step 6.
6. wenn View Quelle ist clean aber GSC still says noindex, don’t trust Ihre browser. ausführen ein live Googlebot fetch (URL Inspection → testen live URL, oder the Rich Results testen). Look für ein CDN/cache serving ein stale header, ein directive conditional auf user-agent, oder ein staging template leaking into production.
7. wenn the URL ist auch disallowed in robots.txt, fix that erste.
ein disallow hides the noindex aus Google entirely, so nothing Sie tun zu the
noindex will nehmen effect until crawling ist allowed again. entfernen the disallow
(oder wait für es zu lift) vor moving auf.
8. entfernen the directive bei its real Quelle — template, server config, CMS field, oder CDN edge rule — und purge any cache sitting in front von es.
9. Re-überprüfen mit ein live Googlebot fetch that the Seite now returns kein noindex, then verwenden Anfrage Indexing für priority URLs and/or Validate Fix für the whole affected festlegen.
10. Wait und recheck. Recrawl und reindexing sind nicht instant, und es gibt kein fixed turnaround — Google’s own guidance ist that revisit timing depends auf the Seite’s importance und kann ausführen zu months für lower-priority Seiten, nicht just days. Anfrage Indexing asks Google zu try sooner; es tut nicht guarantee ein timeline. und wenn Sie disallowed the Seite in step 7 nur zu speichern crawl budget after removal, das ist the one case wo disallow-after-noindex ist correct.
Scripts und snippets
prüfen the response headers (macOS/Linux, shell) — the nur Weg zu sehen ein
X-Robots-Tag, since es never zeigt up in View Quelle:
curl -sI -A "Mozilla/5.0 (compatible; Googlebot/2.1; +http://www.google.com/bot.html)" "https://example.com/page/" | grep -i "x-robots-tag\|^HTTP"prüfen the response headers (Windows, PowerShell) — gleich prüfen, kein curl
erforderlich:
$r = Invoke-WebRequest -Uri "https://example.com/page/" -UserAgent "Mozilla/5.0 (compatible; Googlebot/2.1; +http://www.google.com/bot.html)" -UseBasicParsing
$r.Headers["X-Robots-Tag"]
$r.StatusCodefinden ein meta robots tag in the rendered DOM (DevTools Console) — paste into the Console panel auf the live Seite; catches tags injected durch ein tag manager that raw View Quelle would miss:
[...document.querySelectorAll('meta[name="robots"], meta[name="googlebot"]')].map(m => m.outerHTML)Bookmarklet — prüfen the current Seite’s meta robots tag in one klicken. speichern als ein bookmark mit dies als the URL, then klicken es auf any Seite Sie sind auditing:
javascript:(function(){var m=[...document.querySelectorAll('meta[name="robots"],meta[name="googlebot"]')].map(function(x){return x.outerHTML}).join('\n')||'No meta robots tag found in DOM';alert(m);})();Regex — pull noindex out von ein bulk HTML export. wenn Sie sind grepping ein batch
von saved Seite-Quelle files oder ein crawlen export für content="...noindex..."
Werte, dies captures the full content attribute so Sie kann sehen wenn noindex ist
paired mit anything else (like nofollow oder noarchive):
<meta\s+name=["'](?:robots|googlebot)["']\s+content=["']([^"']*)["']The capture group (([^"']*)) ist the full directive Liste — prüfen es für
noindex specifically anstatt assuming ein match bedeutet noindex, since the gleich
tag kann carry noarchive oder other directives ohne es.
Tools für dies task
dies Website’s Tools
- Google Index Checker — fetches ein URL als
Googlebot und Berichte observable indexability signals: status, redirects,
noindex directives (meta tag und header), und canonical hints, alle in one
pass. es ist explicit that es kann nicht sehen Google’s actual index state — nur
Search Console kann — aber es ist the fastest Weg zu prüfen the noindex + canonical
- status combination vor Sie go anywhere near GSC.
- HTTP Header Checker — raw response headers für ein
URL, einschließlich
X-Robots-Tag. verwenden dies wenn Sie specifically benötigen zu confirm ein header-based noindex (oder confirm es ist gone after ein fix), separate aus anything in the Seite’s HTML. - Robots.txt Tester — prüft whether ein given URL ist disallowed für ein given user-agent. ausführen dies whenever Sie sind diagnosing ein noindex that tut nicht seem zu sein taking effect — ein disallow auf the gleich URL ist the classic cause.
Drittanbieter- Tools
- Google Search Console — the Page Indexing report (wo dies status lives), the sitemap Bericht (für the “Submitted” variant), und URL Inspection → testen live URL, the nur Tool that zeigt Sie ein Echtzeit- Googlebot fetch und verdict.
- Rich Results testen — another live Googlebot fetch; nützlich als ein second lesen auf the HTTP response und rendered snapshot wenn Sie sind chasing ein phantom oder CDN-served noindex.
Validation Tests
ausführen these after removing ein noindex Sie didn’t wollen, oder after resolving ein
“Submitted URL marked ‘noindex’” contradiction.
testen 1: Header kein longer sends X-Robots-Tag: noindex
- testen zu ausführen:
curl -Ithe URL (oder the HTTP Header Checker) und inspect the response headers. - Expected Ergebnis: kein
X-Robots-Tagheader, oder one ohnenoindexin es. - Failure interpretation: The directive ist still being served — prüfen server/CMS config again, und klar any CDN oder edge cache that might sein serving ein stale response.
- Monitoring window: Immediate — dies ist ein live fetch, nicht ein crawlen-dependent signal.
- Rollback trigger: N/A (dies testen tut nicht ändern anything); re-ausführen after jede config oder cache ändern until es passes.
testen 2: Rendered Seite hat kein meta robots noindex
- testen zu ausführen: Google Index Checker oder ein
DevTools/View Quelle prüfen von the rendered
<head>. - Expected Ergebnis: kein
<meta name="robots" content="noindex">(odergooglebotvariant) in the rendered DOM. - Failure interpretation: ein template, tag manager, oder JS injection ist still adding the tag — prüfen rendered DOM, nicht just raw Quelle.
- Monitoring window: Immediate.
- Rollback trigger: N/A; re-ausführen after jede template/config ändern.
testen 3: Live Googlebot fetch confirms the Seite ist indexable
- testen zu ausführen: GSC URL Inspection → testen live URL.
- Expected Ergebnis: “URL ist verfügbar zu Google” mit kein noindex flagged in the live testen Ergebnis.
- Failure interpretation: Googlebot ist seeing something Ihre browser ist nicht — prüfen für ein user-agent-conditional directive oder CDN rule serving Google ein different response than Sie erhalten.
- Monitoring window: Immediate für the live-testen Ergebnis itself.
- Rollback trigger: wenn the live testen still zeigt noindex after ein config ändern plus ein cache purge, treat the fix als nicht yet live und halten debugging vor requesting indexing.
testen 4: URL ist nicht auch blocked by robots.txt
- testen zu ausführen: Robots.txt Tester against the gleich URL.
- Expected Ergebnis: Allowed für Googlebot.
- Failure interpretation: ein disallow ist hiding whatever noindex state exists — Google kann nicht recrawl zu sehen Ihre fix. entfernen the disallow erste.
- Monitoring window: Immediate.
- Rollback trigger: N/A; dies must pass vor the other Tests bedeuten anything für ein Seite Sie wollen indexed.
testen 5: Page Indexing report clears the status
- testen zu ausführen: GSC Page Indexing report, after hitting Validate Fix auf the affected group (oder Anfrage Indexing für ein single priority URL).
- Expected Ergebnis: The URL moves out von “URL marked ‘noindex’” / “Submitted URL marked ‘noindex’” und into “Indexed” (oder Ihre intended status) in the Bericht.
- Failure interpretation: Still pending recrawl, oder the directive ist still present somewhere Sie haven’t checked (re-ausführen Tests 1–3).
- Monitoring window: Variable, nicht fixed — validation läuft in the background auf Google’s own recrawl schedule. Google’s documentation says revisit timing depends auf the Seite’s importance und kann nehmen considerably longer than ein few days, bis zu months für lower-priority Seiten; don’t treat ein slow aktualisieren als ein failure auf its own.
- Rollback trigger: wenn validation still hasn’t moved after ein genuinely long wait (und Tests 1–4 alle pass), re-prüfen Tests 1–4 in order anstatt re-submitting the gleich fix.
Quiz
Five questions zu prüfen war tatsächlich stuck aus dies article.
Änderungsprotokoll
Aktualisiert am 18. Juli 2026.
Redaktionelle Zusammenfassung und aufgezeichnete Änderungsdetails.Änderungsdetails
-
Detaillierte Änderungsangaben sind derzeit auf Englisch verfügbar.
-
Detaillierte Änderungsangaben sind derzeit auf Englisch verfügbar.
-
Detaillierte Änderungsangaben sind derzeit auf Englisch verfügbar.
Vollständiger Vergleich nicht verfügbar — für diese Version wurde kein früherer Schnappschuss archiviert.