SEO technique

A complet guide to SEO technique — a plain-language Beginner's Guide and a systems-level Avancé Guide to exploration, rendering, indexation, and ranking.

Première publication : 25 juin 2026 · Dernière mise à jour : 3 août 2026 · Advanced
Langues

Two guides in un. The Beginner's Guide explique SEO technique from scratch — the explorer → index → rank pipeline, the handful of foundations every site nécessite, Comment vérifier votre propre site, and qui myths to ignore. The Avancé Guide goes systems-deep: budget d’exploration, the rendering decision, canonicalization's ~40 signals, maillage interne, Core Web Vitals as three separate problems, ongoing monitoring, migrations, and AI search. The throughline is the un I toujours come back to — SEO technique is the la plupart important partie of SEO jusqu’à it isn't. It's the foundation que lets content and liens rank, pas a ranking trick of its propre. Vous pouvez't rank une page Google won't index, so the highest-value fonctionner is usually the la plupart boring.

TL;DR — SEO technique is the même explorer → render → index → serve pipeline on every site — there’s aucun separate “technical SEO algorithm” — and it’s a foundation, pas a ranking factor of its propre. Treat the pipeline as a series of gates and diagnose qui un une page is stuck at avant modification anything. The leverage is mostly negative (pas losing ce que you’ve earned), so the boring structural fonctionner — canonicalization, redirections, lien internes — pays meilleur, and it compounds at scale. La plupart sites don’t besoin to manage budget d’exploration; rendering is a separate step que peut lag; Core Web Vitals are three distinct problems and a minor ranking lever; canonicalization is a weighted decision à travers ~40 signals; and from 2025 on, AI search gates eligibility on clean technical signals avant it ranks or cites vous. The highest skill ici is prioritization — knowing ce que to ignore.

SEO technique decides eligibility, pas position

SEO technique is the un partie of SEO whose payoff is almost entirely negative: its job is to garder vous from losing rankings, pas to win les. Google doesn’t hand out positions pour clean plumbing. The même exploration, indexation, and ranking systems run si votre site is pristine or a disaster — there’s aucun separate “SEO technique algorithm” sitting behind les. Ce que SEO technique en réalité decides is si votre pages peut enter ceux systems at tout, and si the engine understands les correctement une fois they’re in.

So the correct mental model isn’t “do technical SEO to rank.” It’s “do SEO technique so votre content and liens are allowed to rank.” Que inversion is the whole raison the unglamorous fonctionner — canonicalization, redirections, lien internes — is the highest-value fonctionner, and pourquoi the unique la plupart utile skill in ce discipline is prioritization: knowing ce que to fix and, simplement as souvent, ce que to leave alone.

The pipeline, as gates

Everything hangs off un pipeline, and the operative clause from Google is “pas tout pages faire it via chaque stage.” Evidence for this claim Google Search describes crawling, indexing, and serving as three stages; discovery is part of the crawling stage, and not every page advances through each stage. Scope: Google Search documentation; conceptual explanation, not a promise of ranking outcomes. Confidence: high · Verified: Google: How Search Works Don’t picture a conveyor belt que carries every page to the finish. Picture a series of gates, chaque with its propre réussir/échouer:

  • Explorer — discovery (liens + sitemaps + push protocols) plus the récupérer. Une page nothing liens to, or un disallowed in robots.txt, may jamais arrive.
  • Render — Google runs votre JavaScript in a recent headless Chrome (the Web Rendering Service) avant it peut entièrement comprendre lune page. Ce is a separate step from the récupérer, it’s stateless, and it peut lag.
  • Index — the engine processes lune page, picks a canonical among duplicates, and decides si to store it. “Indexing isn’t guaranteed” même quand exploration and rendering succeed.
  • Serve — requête understanding, alors ranking à travers nombreux automated systems, alors the search fonctionnalités layered on top.

Garder explorer ≠ render ≠ index ≠ rank separate in votre head and la plupart of SEO technique arrête being mysterious. Quand une page underperforms, vous don’t guess and vous don’t modifier ten choses — vous trouver qui gate it failed and fix que stage.

Un honest caveat avant vous treat quelconque pipeline description as gospel, notamment mine: it’s a model, pas le code source. How Search Fonctionne is a talk I give at conferences que walks via ce whole pipeline (slides on SlideShare), and I ouvrir it with a warning I’ll repeat ici: “ce is my understanding of systems… pas going to be 100% complet or accurate.” Hold it loosely and utiliser it to raison à propos de problems.

Who en réalité fait the exploration

“Googlebot” sounds comme un program. It’s a family — desktop, mobile (the un que matters, since indexation is mobile-first), image, news, video, and ads robots d’exploration — tout drawing from the même budget d’exploration pool, qui is pourquoi a runaway image or parameter explorer peut starve the exploration of votre contenu réel.

And it’s pas simplement moteur de recherches anymore. Quand I analyzed Cloudflare Radar explorer données (an Ahrefs piece I wrote on the nouveau wave of bots), search-engine robots d’exploration encore crawled the la plupart, but AI bots were firmly in second placer and on track to overtake les. Si vous lire votre logs, the cast of characters has modifié — and managing it (qui AI robots d’exploration vous autoriser, and confirming the ones hitting vous are who ils claim) is now partie of the job.

Budget d’exploration: quand it matters, and quand it doesn’t

Google defines budget d’exploration as “the définir of URLs que Google peut and veut to explorer,” définir by explorer capacity (votre serveur’s health) and demande d’exploration (popularity and staleness). Vous raise effective budget two façons: give bots plus capacity, or — far plus souvent — arrêter wasting it. Consolidate duplicates, block low-value spaces, retourner 404/410 pour permanently gone pages, fix soft 404s, garder sitemaps current with accurate lastmod, and éviter long chaîne de redirectionss.

The reassuring partie, and I’ll garder saying it: la plupart sites don’t besoin to worry à propos de budget d’exploration. Google itself indique vous que si votre pages are généralement crawled the même day they’re publié, “you don’t need to read this guide.” It starts to bite autour 1M+ pages, or 10k+ pages que modifier rapidly. Ci-dessous que, spend votre energy elsewhere. Evidence for this claim Google says crawl-budget guidance is mainly relevant to very large sites, including sites with over one million unique pages or over ten thousand rapidly changing pages. Scope: Google Search guidance; the page-count examples are diagnostic starting points, not hard eligibility thresholds. Confidence: high · Verified: Google: Large site crawl budget guide Bing’s Fabrice Canel frames the même idea plus bluntly: moins is plus — fewer URLs to explorer is meilleur pour le SEO.

robots.txt: explorer contrôler, pas index contrôler

The la plupart important distinction in ce whole fichier: robots.txt contrôle exploration, pas indexation. Disallowing une URL arrête bots from fetching it — it fait pas garder it out of the index. A disallowed page peut encore be indexé (URL seulement, aucun content) si autre pages lien to it, and worse, si vous disallow une page vous aussi arrêter Google from ever seeing a noindex tag on it.

So the rules are:

  • Vouloir une page gone from search? Autoriser exploration and ajouter noindex. Jamais utiliser robots.txt to deindex.
  • Vouloir bots to skip a low-value URL space (internal search, infinite faceted combinations) and don’t care à propos de indexation? robots.txt disallow is correct.
  • Managing AI robots d’exploration? Ce is aussi où vous autoriser or block GPTBot, ClaudeBot, PerplexityBot, CCBot and friends — a strategic decision, pas a par défaut.

Canonicalization: a weighted decision, pas a command

Canonicalization is où a lot of avancé SEO technique lives, and it’s widely misunderstood. rel="canonical" is a hint, pas a directive. Google weighs it contre nombreux autre signals — redirections, lien internes, sitemap inclusion, HTTPS, Structure d’URL — quand it picks the representative URL. My deep dive on canonicalization puts the count autour 40 signals que feed canonical selection, qui is pourquoi vous parfois voir “Duplicate, Google chose different canonical than user” in Search Console: votre tag was outvoted.

The practical implications:

  • Don’t send conflicting signals. I spent années on enterprise sites (I ran SEO technique in-house at IBM), and in a talk I give appelé Enterprise SEO Chaos I montrer réel pages que “redirigé to un version, canonicaled to a second, and internally lié to a third.” Pick un URL and faire every signal agree.
  • Signal strength roughly ranks redirection > rel="canonical" > internal liens > sitemap. A 301 is a beaucoup stronger statement que a balise canonical.
  • Contenu dupliqué isn’t a penalty. Google’s Gary Illyes has said roughly 60% of the web is contenu dupliqué, and Google treats some of it as normal — pas a spam violation. The cost is split signals and wasted exploration, pas a punishment. The fix is consolidation, pas panic.

And a remarque on JavaScript: I une fois ran a tester — injecting a rel="canonical" via JavaScript on une page que had none in the HTML — and Google honored it, même though it had publicly said it wouldn’t. Après que surfaced, Google mis à jour its JavaScript SEO documentation. The lesson isn’t “use JS canonicals”; it’s que ce stuff is testable, and the docs aren’t toujours the dernier word.

The rendering decision

Rendering is the step la plupart overviews skip, and it’s où JavaScript sites obtenir into trouble. “During the explorer, Google renders lune page and runs quelconque JavaScript it trouve en utilisant a recent version of Chrome.” Evidence for this claim Google processes JavaScript pages in crawling, rendering, and indexing phases and uses a recent version of Chrome for rendering. Scope: Google Search JavaScript processing; rendering and indexing remain subject to technical and quality constraints. Confidence: high · Verified: Google: JavaScript SEO basics It’s a separate, stateless service, it peut cache resources pour weeks, and it peut lag the initial récupérer — so a JS-dependent modifier peut prendre a pendant que to be reflected.

JavaScript isn’t the enemy ici. As I put it in my guide to JavaScript SEO, JavaScript n’est pas bad pour le SEO, and it’s pas evil — it’s simplement différent from ce que nombreux SEOs are utilisé to. The réel decision is how vous render:

  • Rendu côté serveur (SSR) — safest pour le SEO; the HTML arrives complet.
  • Static generation (SSG/pre-rendering) — meilleur of les deux worlds pour content que doesn’t modifier per requête.
  • Rendu côté client (CSR) — highest risk; le contenu seulement exists après JS runs, so you’re betting on the render step.
  • Dynamic rendering — Google calls it a workaround, pas a recommendation; Bing is plus favorable. Treat it as a bridge, pas a destination.

Two traps to know cold. Premier, lazy-loading: Googlebot doesn’t scroll or click, so content que seulement loads on interaction peut stay invisible — assurez-vous it loads quand it’s in the viewport. Second, liens: Google peut seulement follow a lien that’s a réel <a href> element. A routerLink or a click handler with aucun href n’est pas a crawlable lien. Vérifier the rendered output contre the raw HTML with l’URL Inspection outil whenever vous suspect a gap.

Architecture du site and maillage interne

Lien internes do three jobs at une fois: ils aider bots découvrir pages, ils distribute PageRank, and ils réussir topical context via anchor text. John Mueller has appelé maillage interne “super critical for SEO” and un of the biggest levers vous have on votre propre site — and I agree. It’s un of the highest-ROI choses vous contrôler directement.

A few systems-level points:

  • Orphan pages — pages nothing liens to — are the premier chose to hunt pour. Si it’s pas lié, it’s barely discoverable and obtient almost aucun equity.
  • Architecture is crawl-funnel management. Pages importantes belong fermer to the home page; deep, click-distant pages obtenir crawled moins and rank worse.
  • PageRank sculpting with nofollow is dead (since 2009). Nofollowing internal liens rend que equity evaporate plutôt que redistribute. Manage flow with réel architecture, pas nofollow tricks.

Core Web Vitals: three problems, pas un

The biggest practitioner error with page experience is treating it as a unique “make the site faster” problem. Core Web Vitals are three distinct problems with différent root causes and différent fixes:

  • LCP (Plus grand affichage de contenu) — chargement. Driven by server réponse temps, ressources qui bloquent le rendu, and how fast the principal content asset loads. Target sous 2,5 seconds.
  • INP (Interaction jusqu’au prochain affichage) — interactivity. Driven by JavaScript execution blocking the principal thread. Target sous 200 milliseconds. (INP replaced FID in 2024 — si vous encore voir FID anywhere, the advice is stale.)
  • CLS (Décalage cumulatif de mise en page) — visual stability. Driven by images sans dimensions, late-loading fonts, and injected content. Target sous 0,1.

Two choses matter au-delà the definitions. Field données, pas lab données: Google ranks on real-user CrUX données, pas votre Lighthouse score, so a Lighthouse 65 with bon field données beats a Lighthouse 100 with bad field données. And proportion: I’ll be honest — I don’t think Core Web Vitals have beaucoup impact on SEO, and unless a site is extremely slow I généralement won’t prioritize fixing les pour rankings. Do the fonctionner pour utilisateurs and conversions; simplement don’t oversell it as a ranking lever.

Données structurées: signals pour search and AI

Données structurées (utiliser JSON-LD) doesn’t rank vous, but it rend pages eligible pour résultats enrichis and increasingly helps AI systems parse votre content pour citation. It’s genuinely utile — and genuinely oversold as a ranking signal. My honest framing: la plupart of SEO is doing the basics bien, and content and liens déplacer the needle plus que schema fait. Implement it où it unlocks a rich result or clarifies an entity; don’t expect it to lift rankings on its propre. (And remarque: balisage de données structurées URLs ne sont pas crawlable lien internes — Mueller has confirmed ce.)

International, briefly

Si vous serve multiple languages or regions, utiliser distinct URLs per version and hreflang annotations to map les, and préférer ccTLDs or subdirectories over URL parameters. Don’t auto-redirect by IP — Google explicitly warns contre it and it breaks exploration. International SEO is deep suffisant to be its propre pillar; ce is simplement the technical handshake.

SEO technique is an ongoing system, pas a one-time audit

The framing every competitor guide obtient incorrect: SEO technique n’est pas a checklist vous complet une fois. Sites modifier constantly — deploys break balise canonicals, a release slips a noindex into a template, a nouveau ad script tanks INP, chaîne de redirectionss accumulate. The mature pratique is monitoring and regression detection:

  • Watch GSC Page Indexation pour sudden changements in indexé counts and excluded statuses.
  • Watch Statistiques d’exploration and votre logs pour response-code spikes and crawl-pattern shifts.
  • Re-validate explorer, render, and redirections après every significant deployment.

On log fichiers specifically: I utilisé to treat les as a once-every-few-years troubleshooting outil. That’s modifié. Logs are now the clearest placer to voir qui AI robots d’exploration are en réalité hitting vous and how souvent — something aucun autre outil montre vous as directement — so pour anyone who cares à propos de AI search, they’ve become a lot plus utile que ils were.

Migration de sites: the highest-stakes event

A migration — nouveau domain, HTTP to HTTPS, a replatform, une URL restructure — is the unique highest-stakes technical event, parce que it touches every URL at une fois. Map old to nouveau 1:1, utiliser 301/308 redirection permanentes, garder les in placer indefinitely (I voudrait pas rush to supprimer les — a couple of redirection hops n’est pashing to worry à propos de), and utiliser the GSC Modifier of Adresse outil où it s’applique. Migrations peut be complex and involve a lot of personnes, but don’t panic — vous pouvez fix almost anything que goes incorrect. There’s a complet Migration de sites cluster sous ce pillar.

The modern shift, and it cuts contre the lazy “technical SEO is dead” prendre: from 2025 on, AI search systems decide eligibility avant ils ever rank or cite. To be quoted in an AI réponse votre page généralement has to be cleanly canonicalized, fast suffisant, renderable sans heroics, and structured suffisant to be parsed with confidence. Messy signals don’t simplement lower a ranking now — ils peut supprimer vous from the réponse entirely. Parce que Bing’s index feeds nombreux LLM réponses, Bing Webmaster Outils and IndexNow matter plus que Bing’s search share suggests. Technical hygiene matters plus in the AI era, pas moins.

Où the leverage en réalité is

Si vous prendre un chose from ce guide, faire it prioritization. Spend votre temps on indexation, canonicalization, lien internes, and clean migrations — the fonctionner que decides si pages exist in search and consolidate leur equity. Don’t lose sleep over budget d’exploration, Core Web Vitals, contenu dupliqué, or short chaîne de redirectionss unless vous have a spécifique, diagnosed problem. And don’t chase perfection — I doubt there’s a major website that’s technically perfect, and si là were, I’d worry ils were wasting resources on choses que don’t matter au lieu de choses que do.

Ce hub maps the rest of the pillar: How Search Fonctionne, Migration de sites, On-Page, Moteur de recherche Outils, and JavaScript SEO. Commencer wherever votre site is breaking — the pipeline indique vous qui gate to regarder at premier.

Add an expert note

Pin an expert quote

New person? Create their unclaimed profile at /admin/experts/ → Pin a quote first.