Panduan URL Marked 'noindex' (GSC Status)

What Google Search Console "URL marked 'noindex'" _(terjemahan)_ “URL marked 'noindex'” status berarti — dan "Submitted URL marked 'noindex'" _(terjemahan)_ “Submitted URL marked 'noindex'” variant dan legacy "Excluded by 'noindex' tag" _(terjemahan)_ “Excluded oleh 'noindex' tag” name. When ini adalah intentional vs. sebuah mistake, where noindex lives (meta tag vs. X-Robots-Tag header), robots.txt conflict, phantom/CDN noindex, dan cara fix dan validate.

Pertama kali diterbitkan: 23 Jun 2026 · Terakhir diperbarui: 3 Agu 2026 · Advanced
Bahasa
1 sinyal bukti di halaman ini

"URL marked 'noindex'" _(terjemahan)_ “URL marked 'noindex'” adalah Google Search Console halaman pengindeksan status untuk sebuah halaman Google di-crawl dan ditemukan sebuah noindex directive pada (sebuah meta robots tag atau sebuah X-Robots-Tag header), so ini dipertahankan ini out dari indeks. sama condition, multiple names: saat ini "URL marked 'noindex'" _(terjemahan)_ “URL marked 'noindex'”, sharper sitemap-submitted wording "Submitted URL marked 'noindex'" _(terjemahan)_ “Submitted URL marked 'noindex'” itu reports dan alat commonly gunakan, dan legacy "Excluded by 'noindex' tag." _(terjemahan)_ “Excluded oleh 'noindex' tag.” ini adalah biasanya intentional dan fine — validate URL list sebelum Anda "fix" _(terjemahan)_ “fix” anything. nyata red flag adalah sebuah noindexed halaman masih sitting di Anda sitemap — sitemap submission adalah sebuah hint untuk Google, not sebuah guarantee, dan two directives contradict setiap lainnya. Because noindex adalah crawl-dependent, don't pair ini dengan sebuah robots.txt disallow — Google dapat't see sebuah noindex ini dapat't crawl. periksa two places ( meta tag dan X-Robots-Tag header), debug phantom/CDN noindex dengan sebuah live Googlebot fetch, lalu hapus directive, Validate Fix, dan expect reprocessing untuk take longer daripada sebuah day atau two.

TL;DR — “URL marked ‘noindex’” (terjemahan) “URL marked ‘noindex’” adalah GSC halaman pengindeksan status untuk sebuah halaman Google di-crawl dan ditemukan sebuah noindex pada — meta robots tag atau X-Robots-Tag header, dan Google juga honors sebuah robots meta tag placed di halaman body, not hanya <head>. Multiple names, one state: saat ini “URL marked ‘noindex’,” (terjemahan) “URL marked ‘noindex’,” sharper sitemap-submitted wording “Submitted URL marked ‘noindex’” (terjemahan) “Submitted URL marked ‘noindex’” itu reports dan alat commonly gunakan, dan legacy “Excluded by ‘noindex’ tag.” (terjemahan) “Excluded oleh ‘noindex’ tag.” ini adalah distinct dari robots.txt-blocked (tidak pernah di-crawl) dan dari “Crawled — currently not indexed” (terjemahan) “di-crawl — currently not terindeks” (no directive). biasanya intentional — validate URL list pertama. sebuah noindexed URL masih sitting di Anda sitemap adalah nyata flag (sitemap submission adalah sebuah hint, not sebuah guarantee). noindex adalah crawl-dependent: pair ini dengan sebuah robots.txt disallow dan Google dapat’t see ini, so halaman dapat stay terindeks. When aturan conflict, Google applies more restrictive one. periksa two sources ( rendered HTML dan header HTTP), debug phantom/CDN noindex dengan sebuah live Googlebot fetch (pemeriksaan URL / Rich hasil Test), lalu hapus directive, Validate Fix, dan expect reprocessing untuk take longer daripada sebuah day atau two — Google says ini dapat run untuk months untuk lower-priority halaman.

What status actually berarti

report describes Google’s observed directive, not why sebuah CMS, template, atau CDN ditambahkan ini. Evidence for this claim Google reports URL marked noindex when it encounters a noindex directive and does not index the page. Scope: Google Search Console report terminology and documented Search behavior; the label alone may not prove the underlying root cause. Confidence: high · Verified: Google: Page indexing report Google harus menjadi able untuk crawl halaman untuk observe dan apply noindex. Evidence for this claim Google supports noindex through a robots meta tag or X-Robots-Tag header and must crawl the page to observe it. Scope: Google Search Console report terminology and documented Search behavior; the label alone may not prove the underlying root cause. Confidence: high · Verified: Google: Block indexing with noindex

Google’s own definition adalah precise: when Google tried untuk indeks halaman, ini encountered sebuah noindex directive dan therefore melakukan not indeks ini. key kata adalah encountered — Google memiliki untuk crawl halaman untuk see directive. So ini status carries two facts di once: Google reached halaman, dan halaman told ini not untuk menjadi terindeks.

Evidence for this claim Google reports URL marked noindex when it encountered a noindex directive while trying to index the URL and therefore did not index it. Scope: verified Search Console properties Confidence: high · Verified: Page indexing report

itu’s whole accuracy spine here, dan ini adalah what separates ini status dari -nya neighbors:

  • Robots.txt-blocked → Google adalah tidak pernah allowed untuk crawl, so ini didn’t read apa pun konten atau apa pun directive.
  • “Crawled — currently not indexed” (terjemahan) “di-crawl — currently not terindeks” → Google di-crawl, ditemukan no directive, dan chose not untuk indeks anyway.
  • “URL marked ‘noindex’” (terjemahan) “URL marked ‘noindex’” → Google di-crawl, ditemukan sebuah noindex, dan obeyed ini.
Evidence for this claim Google reports URL marked noindex when it encounters a noindex directive and does not index the page. Scope: Google Search Console report terminology and documented Search behavior; the label alone may not prove the underlying root cause. Confidence: high · Verified: Google: Page indexing report

three names adalah one condition

ini trips people up because label memiliki changed di atas time dan shifts berdasarkan how URL adalah ditemukan:

  • “URL marked ‘noindex’” (terjemahan) “URL marked ‘noindex’” — saat ini, umum label di halaman pengindeksan report. Lives di bawah “Not indexed” (terjemahan) “Not terindeks” (formerly “Excluded” (terjemahan) “Excluded”). ini adalah wording Google’s own saat ini halaman pengindeksan report documentation menggunakan dan defines.
  • “Submitted URL marked ‘noindex’” (terjemahan) “Submitted URL marked ‘noindex’” — yang sama underlying condition, tetapi untuk sebuah URL itu’s juga di sebuah sitemap Anda submitted. ini adalah sharper wording practitioners dan ketiga-party SEO alat commonly report untuk itu combination. Google’s saat ini help documentation doesn’t spell ini out sebagai sebuah separately defined status distinct dari “URL marked ‘noindex,’” (terjemahan) “URL marked ‘noindex,’” so treat exact label sebagai report-dependent — tetapi substance holds regardless dari what sebuah given alat panggilan ini: sebuah sitemap adalah dimaksudkan untuk list URLs Anda ingin di Search (dan submitting one adalah sebuah hint untuk Google, not sebuah guarantee dari pengindeksan), so sebuah noindexed URL sitting di ini adalah sebuah contradiction worth resolving.
  • “Excluded by ‘noindex’ tag” (terjemahan) “Excluded oleh ‘noindex’ tag” — legacy name dari pre-2021 “Index Coverage” (terjemahan) “indeks Coverage” report. masih paling-searched colloquial versi. sama underlying thing.

jika Anda’ve landed here dari apa pun dari itu three, Anda’re di yang sama place.

adalah ini sebuah masalah? intentional-vs-accidental decision

Don’t reflexively “fix” (terjemahan) “fix” ini. decision tree:

  1. Pull list dari affected URLs (click ke status).
  2. adalah ini halaman Anda dimaksudkan untuk exclude? Thank-Anda halaman, internal search, faceted/filter URLs, account/admin, staging itu shouldn’t menjadi live. → No tindakan. “Not indexed” (terjemahan) “Not terindeks” adalah not yang sama sebagai “broken,” (terjemahan) “broken,” dan Google says sebagai much: ini URLs memiliki not telah terindeks, tetapi not necessarily karena sebuah error.
  3. adalah sebuah halaman Anda ingin terindeks di ini list? → sebuah noindex leaked onto ini. temukan dan hapus ini.
  4. adalah noindexed URL juga sitting di Anda submitted sitemap (sering surfaced sebagai “Submitted URL marked ‘noindex’” (terjemahan) “Submitted URL marked ‘noindex’”)? → Resolve contradiction directly: hapus noindex (untuk indeks ini) atau hapus URL dari Anda sitemap (untuk leave ini noindexed). Don’t leave sebuah noindexed URL sitting di sebuah sitemap — sitemap submission adalah sebuah hint untuk Google tentang what Anda ingin terindeks, not sebuah permintaan itu overrides halaman’s own directive.

alasan competitors treat ini sebagai sebuah pure “error to fix” (terjemahan) “error untuk fix” adalah itu mereka skip langkah 2. sebagian besar dari time, ini status adalah sistem berfungsi correctly.

Where noindex lives: meta tag vs. X-Robots-Tag header

ada exactly two delivery metode, dan Anda memiliki untuk periksa both because mereka look completely berbeda:

  • Meta robots tag<meta name="robots" content="noindex"> di halaman’s <head>. Targets semua crawler; <meta name="googlebot" content="noindex"> targets Google hanya. ini adalah one Anda dapat spot di HTML.
  • X-Robots-Tag header HTTPX-Robots-Tag: noindex di server’s respons headers. ini adalah sneaky one. ini adalah set di server, CMS, atau CDN config, not di halaman source, so “View Source” (terjemahan) “View Source” won’t tampilkan ini. header metode adalah juga satu-satunya cara untuk noindex non-HTML files — sebuah respons header dapat menjadi digunakan untuk non-HTML resources such sebagai PDFs, video files, dan image files, which memiliki no <head> untuk hold sebuah meta tag.
Evidence for this claim Google supports noindex through a robots meta tag or an X-Robots-Tag HTTP response header; both have the same effect, and the response header supports non-HTML resources. Scope: HTML and non-HTML web resources Confidence: high · Verified: Block Search indexing with noindex

When GSC says noindex dan Anda swear halaman doesn’t memiliki one, header adalah pertama place untuk look ( cheat sheet tab lays two side oleh side).

cara temukan directive pada sebuah halaman

  1. View Source / rendered DOM — search untuk noindex. periksa rendered <head>, not hanya raw source, since sebuah tag dapat menjadi injected oleh JavaScript atau sebuah tag manager. Don’t stop di <head>, either: Google memiliki said ini doesn’t enforce meta-robots placement dan respects sebuah robots meta tag ditemukan di halaman’s <body> too, so sebuah directive injected lower di document masih counts.
  2. respons headerscurl -I https://example.com/page/ (atau browser DevTools → Network → document permintaan → respons Headers) dan cari sebuah X-Robots-Tag line.
  3. pemeriksaan URL (GSC)Test live URL — ini fetches halaman sebagai Googlebot dan reports pengindeksan verdict dan respons. ini adalah one itu catches directives disajikan hanya untuk Google.
  4. Rich hasil Test — lainnya nyata Googlebot fetch itu mengembalikan HTTP respons dan sebuah rendered snapshot dari exactly what server menampilkan Google.
  5. periksa untuk conflicting robots aturan, not hanya sebuah single tag. jika more daripada one robots directive applies untuk halaman (say, sebuah template sets index tetapi sebuah plugin atau header menambahkan noindex), Google applies more restrictive aturan — so sebuah stray noindex anywhere wins bahkan jika lainnya aturan says index. Don’t stop searching setelah Anda’ve ditemukan one directive itu looks permissive.

robots.txt conflict (why disallow + noindex backfires)

ini adalah single sebagian besar-muddled poin di setiap lainnya guide, so I ingin ini exact. noindex adalah crawl-dependent: Google memiliki untuk menjadi able untuk fetch halaman untuk read directive. Google states aturan plainly — untuk noindex aturan untuk menjadi effective, halaman harus not menjadi blocked oleh sebuah robots.txt file dan memiliki untuk menjadi otherwise accessible untuk crawler; jika ini adalah blocked atau crawler dapat’t access ini, crawler akan tidak pernah see noindex, dan halaman dapat masih appear di search (misalnya, jika lainnya halaman tautan untuk ini).

Evidence for this claim Google must be allowed to crawl and otherwise access a URL to see and apply its noindex rule; a robots.txt block can hide the directive while the URL remains eligible to appear from other information. Scope: HTML and non-HTML web resources Confidence: high · Verified: Block Search indexing with noindex

I’ve written tentang flip side dari ini untuk years. di my Ahrefs piece pada “Indexed, though blocked by robots.txt” (terjemahan) “terindeks, though blocked oleh robots.txt”, core poin adalah itu “crawling and indexing are two different things” (terjemahan) “crawling dan pengindeksan adalah two berbeda things” — “if you block a page from being crawled, Google may still index it.” (terjemahan) “jika Anda block sebuah halaman dari menjadi di-crawl, Google dapat masih indeks ini.” dan specifically pada ini conflict: “Unless Google can crawl a page, they won’t see the noindex meta tag and may still index it because it has links.” (terjemahan) “Unless Google dapat crawl sebuah halaman, mereka won’t see noindex meta tag dan dapat masih indeks ini because ini memiliki tautan.” So self-defeating combo adalah noindex + robots.txt disallow: disallow hides noindex, dan halaman dapat stay terindeks via tautan eksternal.

correct sequence untuk actually hapus sebuah halaman:

  1. Allow crawling dan pertahankan noindex di place.
  2. Wait untuk Google untuk recrawl, see directive, dan drop halaman.
  3. hanya lalu, jika Anda ingin save anggaran crawling, Anda dapat disallow ini di robots.txt — setelah deindexing, not sebelum.

My standing recommendation, dari itu sama artikel: “Just add a noindex meta robots tag and make sure to allow crawling — assuming it’s canonical.” (terjemahan) “hanya tambahkan sebuah noindex meta robots tag dan pastikan untuk allow crawling — assuming ini adalah canonical.”

Phantom noindex: CDN cache dan Googlebot-hanya directives

hardest versi dari ini adalah “phantom” (terjemahan) “phantom” noindex: Anda lihat halaman, see no noindex anywhere, dan GSC masih reports one. John Mueller memiliki addressed exactly ini — di cases he’s seen, there adalah sebuah actual noindex, sometimes ditampilkan hanya untuk Google, which dapat menjadi very hard untuk debug. (He noted itu scenario when ini came up; I’m paraphrasing his poin alih-alih quoting ini sebagai sebuah formal statement.)

usual suspects — treat ini sebagai hypotheses untuk periksa, not confirmed causes, until live respons actually menampilkan one dari them:

  • sebuah CDN atau cache serving sebuah stale X-Robots-Tag: noindex header itu’s no longer di Anda origin config.
  • sebuah directive conditional pada pengguna-agent — server mengembalikan sebuah clean halaman untuk Anda browser dan sebuah noindexed one untuk Googlebot.
  • sebuah staging/template leak — sebuah noindex dimaksudkan untuk sebuah staging environment shipping untuk production melalui sebuah shared template.

diagnosis adalah yang sama di semua three: don’t trust “View Source” (terjemahan) “View Source” di Anda own browser. melakukan sebuah nyata Googlebot fetch — pemeriksaan URL’s Test live URL atau Rich hasil Test — which menampilkan Anda respons HTTP dan rendered halaman exactly sebagai Google menerima ini. itu’s how Anda catch sebuah server/CDN serving one thing untuk Anda dan lainnya untuk crawler.

cara fix ini dan validate

setelah Anda’ve confirmed noindex adalah sebuah mistake:

  1. hapus directive di -nya nyata source — meta tag di template, atau X-Robots-Tag header di server/CMS/CDN config. jelas apa pun CDN/halaman cache so fix adalah actually menjadi disajikan.
  2. Confirm dengan sebuah live Googlebot fetch (pemeriksaan URL → Test live URL) itu halaman now mengembalikan no noindex dan menampilkan “URL is available to Google.” (terjemahan) “URL adalah available untuk Google.”
  3. permintaan pengindeksan untuk tinggi-priority URLs, dan/atau gunakan report’s Validate Fix button untuk tell Google untuk recheck whole affected set.
  4. Wait. Deindexing dan reindexing aren’t instant — Google memiliki untuk recrawl untuk see perubahan pertama, dan Google’s own guidance adalah itu revisit timing depends pada halaman’s importance dan dapat take considerably longer daripada sebuah day atau two (-nya documentation gives “months” (terjemahan) “months” sebagai sebuah possibility untuk lower-priority halaman). permintaan pengindeksan pada sebuah priority URL adalah how Anda tanyakan Google untuk try sooner, not sebuah cara untuk force sebuah spesifik timeline. Don’t panic jika status lingers while reprocessing.

untuk inverse — sebuah halaman Anda ingin noindexed tetapi itu’s stuck di indeks because ini adalah juga robots.txt-blocked — unblock crawling pertama so Google dapat finally see noindex.

Where ini sits

ini status adalah one node di Google’s halaman pengindeksan report. robots directive behind ini — noindex — dapat menjadi delivered sebagai sebuah meta robots tag atau sebuah X-Robots-Tag header, dan right alat depends pada whether Anda’re berfungsi dengan HTML atau non-HTML files. neighboring statuses (“Indexed, though blocked by robots.txt” (terjemahan) “terindeks, though blocked oleh robots.txt” dan “Crawled — currently not indexed” (terjemahan) “di-crawl — currently not terindeks”) describe berbeda states dan perlu berbeda fixes; keeping them straight adalah sebagian besar dari battle.

Add an expert note

Pin an expert quote

New person? Create their unclaimed profile at /admin/experts/ → Pin a quote first.