HTTP kody błędów

How HTTP 4xx i 5xx kody błędów affect SEO — how Google handles błędy, który ones cause deindexing, crawl waste, i ranking drops, i how to monitorować i fix them.

Opublikowano po raz pierwszy: 27 cze 2026 · Ostatnia aktualizacja: 3 sie 2026 · Advanced
Języki
1 sygnał dowodowy na tej stronie

HTTP kody błędów są the 4xx (client błąd) i 5xx (serwer błąd) odpowiedzi a serwer zwroty zamiast a udany 2xx. Google treats the two classes bardzo differently: 4xx (except 429) means 'the treść doesn't exist' — the strona drops z the index, z zero effect on częstotliwość indeksowania; 5xx (i 429) means 'the serwer jest failing' — Google throttles crawling witryna-wide pierwszy, i tylko drops the strony if the błędy persist. 429 jest numerically a 4xx ale Google explicitly treats it as a serwer błąd. A soft 404 jest the trap case — a 200 że reads as 'nie znaleziono,' który Google flags as wasting budżet indeksowania (mainly a duży-witryna concern) zamiast cleanly dropping the strona. 404s są nie a quality lub ranking signal (Mueller), so don't panic ponad an błąd count — an isolated błąd's rzeczywisty impact nadal depends on co the URL jest (a special zasób like robots.txt gets jego own błąd handling, unlike an ordinary strona) — prioritize the errored URLs że mieć links lub ruch. dla planned downtime używać 503 + Retry-po (nigdy 403/404/410), i nigdy 503 twój robots.txt. monitoruj whole family przez Search Console's strona indeksowanie raport. ten hub maps i links to każdy individual code: 401, 403, 404, 404 vs 410, 410, 429, 451, 500, 502, 503, 504, i soft 404.

TL;DR — HTTP kody błędów są the 4xx (client) i 5xx (serwer) status classes. Google draws a trudny behavioral wiersz między them: 4xx (except 429) means “the content doesn’t exist” — the URL jest dropped z the index, z “no effect on crawl rate”; 5xx (i 429) means “the server is failing” — Google throttles crawling proportionately, preserves zindeksowany URLs at pierwszy, i tylko deindexes if the błędy persist, then ramps częstotliwość indeksowania back up gradually po odzyskiwanie. 429 jest numerically a 4xx ale Google calls it “a server error.” treść z dowolny błąd odpowiedź jest ignored. 404s są nie a quality signal (Mueller) — triage by links/ruch, don’t fix everything, i remember an isolated błąd’s rzeczywisty impact depends on the URL (robots.txt ma special błąd handling że an ordinary strona doesn’t). Soft 404s (a 200 że reads as “not found”) są flagged by Google as wasting budżet indeksowania — mainly a duży-witryna concern, nie a guaranteed effect on każdy witryna. dla planned downtime używać 503 + Retry-po, kept to “a few days at most,” i nigdy 503 the robots.txt. monitoruj family przez GSC’s strona indeksowanie raport. ten hub maps i routes to każdy individual code.

The one distinction że runs the whole topic

The protokół code describes the HTTP outcome; a Search Console label describes how Google classified an observed fetch. Evidence for this claim Primary standard or official documentation supporting the adjacent article claim. Scope: Protocol semantics and Search behavior are kept separate; no indexing, ranking, or migration-timing guarantee is inferred. Confidence: high · Verified: RFC 9110: Status codes robić nie infer a single root cause lub dokładny removal time z the family alone. Evidence for this claim Primary standard or official documentation supporting the adjacent article claim. Scope: Protocol semantics and Search behavior are kept separate; no indexing, ranking, or migration-timing guarantee is inferred. Confidence: high · Verified: Google: HTTP and network errors

Four warstwy get flattened do “it’s a 404” whenever people talk o błędy, i keeping them apart jest co makes the rest of ten strona make sense:

  1. protokół semantics — co the kod stanu means per the HTTP spec (a 404 means “not found,” pełny stop).
  2. The observed fetch — co Googlebot actually got back on a specific żądanie at a specific time, który może differ z co a przeglądarka sees.
  3. Search processing — how Google’s indeksowanie pipeline classifies i acts on że observed fetch (drops the URL, throttles the crawl, ignores the body).
  4. Root cause — the rzeczywisty powód on twój serwer (a bad deploy, an overloaded baza danych, a WAF reguła), który the kod stanu alone nigdy tells you.

Knowing the family (4xx vs 5xx) tells you który of the pierwszy three warstwy you’re looking at. It nigdy substitutes dla warstwa 4 — you nadal mieć to go find out why.

Google’s own docs frame the split cleanly, i it’s worth internalizing przed anything else: 4xx błędy say “the content doesn’t exist”; 5xx błędy say “the server itself is failing.” tamte są two completely różny problems, i Google’s indeksowanie pipeline responds to them in two completely różny ways.

A użyteczny anchor pierwszy: even a sukces code isn’t a promise. Google says plainly że “for Google Search, an HTTP 2xx (success) status code doesn’t guarantee indexing.” kody błędów są actually the więcej deterministic half of the picture — they tell Google, unambiguously, either “gone” lub “broken.”

How Google treats 4xx błędy

dla client błędy, Google’s behavior jest uniform i blunt: “Google doesn’t use the content from URLs that return 4xx status codes,” i “Google crawlers inform the next processing system that the content doesn’t exist.” nawet jeśli twój 403 lub 404 strona renders a wall of rzeczywisty tekst, none of it gets zindeksowany — the body of an błąd odpowiedź jest ignored.

Two consequences follow:

  • Previously zindeksowany URLs get dropped. Once a strona reliably zwroty a 4xx, Google removes it z the index ponad time.
  • There’s no crawl-rate penalty. ten jest the myth-buster: “The 4xx status codes, except 429, have no effect on crawl rate.” A mountain of 404s robi nie slow Google’s crawling of the rest of twój witryna.

że ostatni point kills a bad instinct: don’t try to throttle Googlebot z a 401 lub 403. Google explicitly warns “don’t use 401 and 403 status codes for limiting the crawl rate.” An auth wall doesn’t slow the crawl — it just makes the treść invisible.

How Google treats 5xx błędy (i 429)

serwer błędy trigger the “the server is struggling” branch, i here Google jest deliberately protective of twój witryna:

  • częstotliwość indeksowania throttles pierwszy, proportionate to volume. “5xx and 429 server errors prompt Google’s crawlers to temporarily slow down with crawling,” i “the decrease in crawl rate is proportionate to the number of individual URLs that are returning a server error.” A handful of 500s barely registers; a witryna-wide outage cuts crawling trudny.
  • zindeksowany URLs są preserved… until the błędy persist. “Already indexed URLs are preserved in the index, but eventually dropped,” i Google “removes from the index URLs that persistently return a server error.” A short blip doesn’t deindex you; a sustained awaria robi.
  • treść z 5xx jest ignored too. “Any content Google receives from URLs that return a 5xx status code is ignored.”
  • odzyskiwanie jest automatic ale gradual. “Once the server starts responding with a 2xx status code, Google gradually increases the crawl rate for the site.” Google jest fast to back off i cautious to ramp back up — there’s no manual “unlock,” you just fix the root cause i wait.

Why 429 lives in the serwer-błąd bucket

ten jest the single najbardziej-missed nuance in najbardziej poradniki. 429 Too Many Requests jest numerically a 4xx, ale Google treats it as a serwer signal: “Google’s crawlers treat the 429 status code as a signal that the server is overloaded, and it’s considered a server error.” So a WAF lub rate limiter że starts firing 429s at Googlebot będzie throttle twój crawl the same way a wave of 500s by — nie drop individual strony the way a 404 robi. gdy I break the codes down in my HTTP kody stanu poradnik at Ahrefs, I put 429 z the serwer błędy dla exactly ten powód: it’s “a form of rate-limiting to protect the server,” i it makes Google slow down.

The błąd-code family map

Here’s the fast triage view of the whole family. każdy code below jest jego own deep dive nested poniżej ten hub (they’re in the sidebar too):

Blocked / access błędy

  • 401 Unauthorized — the client hasn’t identified lub verified itself gdy needed. Blocked to Googlebot behind an auth żądanie.
  • 403 Forbidden — the client jest znany ale doesn’t mieć access rights.

nie-found błędy

  • 404 nie znaleziono — the requested zasób isn’t found.
  • 404 vs 410 — the practical difference między “not found” i “gone” (it’s smaller than people think).
  • 410 Gone — like a 404, ale it również says the zasób won’t być back. Drops strony slightly faster.

The dual-tożsamość code

  • 429 zbyt wiele żądań — rate limiting; numerically 4xx ale Google treats it as a serwer błąd dla crawl-rate purposes.

Legal

  • 451 niedostępny z powodów prawnych — blocked dla a legal powód: kraj-level bloki, DMCA takedowns.

serwer błędy (5xx)

  • 500 wewnętrzny błąd serwera — the serwer hit an problem it może’t handle.
  • 502 zła brama — a bad odpowiedź z an upstream serwer.
  • 503 Service Unavailable — the serwer jest overloaded lub down dla maintenance (the poprawny code dla planned downtime).
  • 504 Gateway Timeout — no timely odpowiedź z an upstream serwer.

The trap case

  • Soft 404 — a strona że zwroty 200 OK ale reads as “not found.” The worst of oba worlds — więcej on ten below.

Broken redirects są błąd-adjacent ale categorized osobno (Google surfaces them as “Redirect error” in Search Console, distinct z a normal, działający “Page with redirect”). By domyślny Google’s crawlers “follow up to 10 redirect hops” przed giving up; a chain że’s too long, loops, lub contains a bad URL turns do że Redirect-błąd status.

robi an HTTP błąd hurt twój SEO?

Lead z the verdict: an isolated błąd on an ordinary strona jest almost nigdy a problem, i 404s specifically są nie a ranking lub quality signal. Mueller ma był explicit i repeated o ten. The reflex — “I have 50,000 404s, my site must be penalized” — jest the liczba-one myth to defuse. błędy są a normal part of the web; having strony 404 lub 410 jest the technically poprawny way to handle URLs że don’t exist.

“Isolated” doesn’t mean “always harmless,” though — the rzeczywisty impact depends on co’s returning the błąd. An ordinary strona 404ing jest a non-event; a special zasób jest różny. Google gives robots.txt jego own błąd-handling reguły distinct z regular URLs, so a serwer błąd on robots.txt może affect crawling in a way an ordinary strona’s 404 nigdy by. Judge an błąd by co URL it’s on i co depends on it, nie by the kod stanu alone.

The one pattern z a rzeczywisty, udokumentowany mechanism jest persistent, mass 5xx: crawl-rate throttle → eventual deindexing. i even że jest proportionate to how wiele URLs są erroring i generally reversible once the serwer recovers — Google describes the ramp-back-up as gradual, nie instant, so treat the dokładny timing, retry cadence, i odzyskiwanie speed as dowód-dependent (co Google’s docs describe happening) zamiast a fixed guarantee. The distinction Mueller draws jest the one to remember: częstotliwość indeksowania reacts to serwer błędy (429/500/503/ timeouts), nie to 404s.

Crawl waste vs deindexing — two różny harms

It pomaga to oddzielny the two ways błędy może cost you:

  • Deindexing — persistent 4xx (strona dropped as “gone”) lub persistent 5xx (strony dropped po sustained serwer awaria). ten jest o strony leaving the index.
  • Crawl-budget waste — mostly a duży-witryna concern. Google’s crawl-budget poradnik notes “if the site slows down or responds with server errors, the limit goes down and Google crawls less.” ale the rzeczywisty budget-waster jest the soft 404: “soft 404 pages will continue to be crawled, and waste your budget.” ponieważ a soft 404 looks alive (it zwroty 200), Google zachowuje re-checking a strona że isn’t really there. że’s why the soft 404 jest the “worst of both” case — it doesn’t cleanly drop like a prawdziwy 404, it lingers i burns fetches.

The soft-404 trap ma two common causes, oba worth naming: a custom “page not found” template że zwroty 200 zamiast a rzeczywisty 404, i a blanket redirect of każdy dead URL to the homepage — Google recognizes że pattern as a soft 404 too. A friendly 404 strona jest good dla UX i fully recommended — as long as it nadal zwroty the rzeczywisty HTTP 404 kod stanu.

Doing planned downtime the right way

gdy you intentionally take a witryna (lub a sekcja) offline, the poprawny code jest 503 Service Unavailable z a Retry-After header — nigdy a 4xx. Google’s “Pause your online business” guidance jest specific:

  • “If you need to urgently disable the site for 1-2 days, then return an informational error page with a 503 HTTP response status code.”
  • “This is an extreme measure that should only be taken for a very short period of time (a few days at most),” ponieważ “completely closing a site even for just a few weeks can have negative consequences on Google’s indexing of your site.”
  • “Don’t block the website by returning 403, 404, 410 HTTP status codes” podczas downtime — a 4xx says “permanently gone,” który jest exactly the błędny signal dla a temporary outage.
  • The trap almost everyone misses: “Don’t return a 503 HTTP response status code for the robots.txt file because this blocks all crawling.” 503 twój strony, nie twój robots.txt.

How to monitoruj whole family in Search Console

Day to day, you’ll meet te błędy in Search Console’s strona indeksowanie raport, gdzie każdy maps to a distinct status:

  • nie znaleziono (404)“this page returned a 404 error when requested.”
  • serwer błąd (5xx)“your server returned a 500-level error when the page was requested.”
  • Blocked due to nieautoryzowane żądanie (401)“the page was blocked to Googlebot by a request for authorization.”
  • Blocked due to dostęp zabroniony (403) — a 403 gdzie credentials były provided ale access wasn’t granted.
  • Blocked due to other 4xx problem — a 4xx nie covered by another status; używać URL Inspection to debug.
  • Soft 404 — a “user-friendly ‘not found’ message but not a 404 HTTP response code.”
  • Redirect błąd — chain too long, a loop, an ponad-long URL, lub a bad URL in the chain.

każdy status points at a różny root cause i fix path. Once you’ve resolved one, używać walidować Fix to prompt a recrawl — ale ustawić realistic expectations on timing; recrawl isn’t instant. i remember Google’s own framing: “it’s fine for a URL not to be indexed for the right reasons — for example… a 404 for a page that you’ve removed and have no replacement for.” nie każdy błąd jest a to-robić.

Search Console alone isn’t enough to act on — it’s one vantage point, sampled i delayed. gdy you log an błąd dla triage, record at minimum: the URL, gdy you observed it, the vantage point (Search Console vs. a live sprawdzenie vs. serwer/CDN logs), the użytkownik agent the żądanie came in on, the HTTP metoda, the final kod stanu i odpowiedź path (w tym dowolny łańcuch przekierowań), whether it’s a one-off lub recurring, i a post-fix verification krok once you’ve redeployed. Triangulating GSC wobec a live status sprawdzenie i twój own logs jest co turns a stale raport row do a confirmed, fixable problem.

How to fix i prioritize błędy

  • Triage 404s by wartość. Fix the ones z inbound links, linki wewnętrzne, mapa witryny presence, lub lingering ruch — 301-redirect tamte to a relevant strona to recover the link equity. Let genuinely dead URLs 404 lub 410. As I put it in my Ahrefs poradnik, the practical fix dla najbardziej of te jest że “you just need to 301 redirect each of these pages to a relevant page” — ale tylko gdzie a relevant target exists. Don’t blanket-redirect everything to the homepage (że’s a soft 404).
  • 410 vs 404 jest marginal. A 410 drops a strona slightly faster than a 404; the practical SEO difference jest minor. używać 410 gdy you want to być explicit że something jest gone dla good, ale don’t expect it to być dramatically better.
  • Root-cause 5xx. The fixes są on the serwer side: capacity i timeouts dla 500s, upstream/CDN health dla 502/504, i WAF lub rate-limiting reguły misfiring on Googlebot dla 403/429. Confirm the rzeczywisty bot z a reverse/forward DNS sprawdzenie przed you go rate-limiting it.
  • Common root causes worth naming: app i baza danych błędy i overloaded hosts (5xx), upstream/CDN awarie (502/504), rate limiting lub bot-blocking WAFs (403/429), broken migrations i stale linki wewnętrzne (404), i misconfigured “friendly error pages” (soft 404).

Bing: a similar pattern, nie independently verified code-dla-code

Bing’s public statements point in the same direction as Google’s approach — 400-range codes są treated as missing lub forbidden, i 500-range codes signal serwer trouble że działa wobec crawl efficiency — ale I haven’t independently verified pełny, current, code-by-code parity wobec Bing’s own documentation, so treat ten as directional zamiast a confirmed one-to-one match. Fabrice Canel frames Bing’s goal as a “crawl efficiency north star … to crawl a URL only when the content has been added … updated,” i persistent błędy działać directly wobec że — Bing spends crawl footprint on URLs że aren’t yielding fresh, indeksowalny treść. Bing również recommends a 503 z Retry-po dla planned downtime zamiast serving błąd strony as 200. You’ll find Bing’s crawl-błąd surfaces in Bing narzędzia dla webmasterów (URL Inspection, Crawl Control, witryna Scan).

gdzie to go następny

ten strona jest the conceptual hub dla the błąd-code family. It sits inside the broader HTTP kody stanu cluster (the pełny 1xx–5xx picture, plus redirects i sukces codes); ten sub-hub jest the map dla the błąd half of że. każdy code below jest jego own deep dive:

Blocked / access

  • 401 Unauthorized — co triggers the “Blocked due to unauthorized request” status, i why auth walls don’t throttle Googlebot.
  • 403 Forbidden — credentials-provided-ale-denied, i the WAF/bot-blocking patterns że cause fałszywy 403s to Googlebot.

nie znaleziono

  • 404 nie znaleziono — how Google handles missing strony, why it’s nie a penalty, i który 404s to actually fix.
  • 404 vs 410 — the rzeczywisty, mały difference, i gdy to reach dla każdy.
  • 410 Gone — the “permanently gone” signal i jego slightly faster drop.

Rate limiting

  • 429 zbyt wiele żądań — the 4xx że behaves like a 5xx, i how to zachować rate limits z throttling twój crawl.

Legal

  • 451 niedostępny z powodów prawnych — takedowns, kraj bloki, i how legal removals pokazywać up.

serwer błędy

  • 500 wewnętrzny błąd serwera — the generic serwer awaria i how to root-cause it.
  • 502 zła brama — upstream/proxy awarie.
  • 503 Service Unavailable — the poprawny code dla maintenance i planned downtime (z the robots.txt trap).
  • 504 Gateway Timeout — upstream timeouts.

The trap case

  • Soft 404 — the 200-że-reads-as-gone, why it wastes budżet indeksowania, i how to turn it do a rzeczywisty 404.

Broken redirects są handled osobno as the Redirect błąd status — powiązany ale categorized on jego own in Search Console.

Add an expert note

Pin an expert quote

New person? Create their unclaimed profile at /admin/experts/ → Pin a quote first.