SEO technique Checklist

A comprehensive SEO technique checklist covering crawlability, indexation, Core Web Vitals, données structurées, sitemaps, robots.txt, canonicalization, HTTPS, and mobile SEO — organized by priority and site type.

Première publication : 27 juin 2026 · Dernière mise à jour : 3 août 2026 · Advanced
Langues
1 indice probant sur cette page

A SEO technique checklist confirms the technical conditions a site doit meet pour moteur de recherches and AI réponse engines to explorer, render, index, and serve its pages — Google's four load-bearing conditions are crawlable, indexable, understandable, renderable. It's the technical-only slice; the broader full-site version lives in the SEO Audit Checklist. Fonctionner it in priority order — crawlability, indexation, HTTPS, mobile parity, sitemaps, données structurées, Core Web Vitals — and gate whole sections by site type: la plupart petit sites peut skip crawl-budget engineering entirely, qui Google confirms in its propre docs. Two myths to kill on sight: robots.txt ne fait pas deindex une page, and données structurées n’est pas a ranking factor. Meeting the checklist is necessary but pas sufficient — Google dit indexation encore isn't guaranteed.

TL;DR — SEO technique is the floor, pas the ceiling: pages doit be crawlable → indexable → understandable → renderable avant content fonctionner peut pay off. Run the checklist in priority order and gate sections by site type — la plupart petit sites peut skip budget d’exploration, faceted-nav contrôler, log analysis, and JS-rendering engineering, qui Google’s propre crawl-budget doc confirms. Kill two myths on sight: robots.txt doesn’t deindex, and données structurées isn’t a ranking factor. Treat Core Web Vitals as targets, pas réussir/échouer gates — that’s Google’s propre “strive to” language. And meeting every box encore doesn’t guarantee indexation.

Evidence for this claim Google's minimum technical requirements include accessible Googlebot crawling, a successful HTTP response, and indexable content. Scope: Eligibility prerequisites, not a guarantee of indexing or ranking. Confidence: high · Verified: Google Search Essentials: Technical requirements Evidence for this claim Meeting technical requirements does not guarantee that Google will crawl, index, or serve a page. Scope: Google Search eligibility and selection behavior. Confidence: high · Verified: Google Search Essentials: Technical requirements

The mental model: four conditions

Google frames the whole of SEO technique autour four load-bearing conditions — une page has to be crawlable, indexable, understandable, and renderable. Its Search Essentials technical requirements boil que bas to three minimums: “Googlebot isn’t blocked,” “The page works” (served with an HTTP 200 status), and “The page has indexable content.” Everything on ce checklist is really in service of ceux.

The critical caveat, straight from the même doc: “Simplement parce que une page meets ces requirements doesn’t mean que une page va be indexé; indexation isn’t guaranteed.” SEO technique is a gate vous have to réussir via, pas a lever que guarantees results. Ce is pourquoi I’ve toujours argued the technical checklist is the floor — vous clair it so content and liens peut do leur job, pas au lieu de les.

Avant vous commencer: qui sections même appliquer to vous?

Every competing checklist organizes by topic and runs the whole chose on every site. That’s the incorrect par défaut. The meilleur organizing axis is site type, parce que Google itself gates its la plupart avancé guidance by site size. From the large-site crawl-budget guide: “Si votre site doesn’t have a grand number of pages que modifier rapidly, or si votre pages sembler to be crawled the même day que ils are publié, vous don’t besoin to lire ce guide.”

Google’s propre (deliberately rough) thresholds pour quand budget d’exploration starts to matter:

  • Grand sites — 1 million+ unique pages with content modification à propos de weekly.
  • Medium-or-larger sites — 10 000+ unique pages with very rapidly modification (daily) content.
  • Sites with a big share of URLs stuck in Search Console’s “Découvert - currently non indexée” state.

So here’s the split I’d en réalité run:

Starter track — petit / brochure / local-business sites (< ~10K URLs): crawlability sanity vérifier, indexation/canonical sanity vérifier, HTTPS, mobile parity, un XML sitemap, un round of Core Web Vitals, valid données structurées où it earns a rich result. Skip budget d’exploration, faceted-nav contrôler, log-file analysis, and JS-rendering engineering entirely.

Avancé track — grand / ecommerce / JS-heavy / enterprise sites: everything in the Starter track plus crawl-budget management, faceted-navigation contrôler, JavaScript rendering audits, server-log analysis, and (si international) hreflang.

1. Crawlability

  • robots.txt correctness. Confirmer vous aren’t disallowing anything vous vouloir indexé. Google: “A robots.txt fichier indique moteur de recherche robots d’exploration qui URLs the robot d’exploration peut accès on votre site. Ce is utilisé mainly to éviter overloading votre site with requêtes; it n’est pas a mechanism pour keeping a web page out of Google.” Utiliser it to garder bots out of low-value spaces (internal search, infinite parameter combinations), pas as a deindexing outil. Ce is the même territory the robots.txt deep dive covers in complet.
  • The robots.txt + noindex contradiction trap. Do pas block une URL in robots.txt and rely on a noindex on it. Google: “Pendant que Google won’t explorer or index le contenu blocked by a robots.txt fichier, we pourrait encore trouver and index a disallowed URL si it is lié from autre places on the web.” The bot can’t explorer lune page, so it jamais sees the noindex, and l’URL peut encore surface bare in results. To supprimer une page: autoriser exploration + noindex, or password-protect it.
  • Explorer errors. Fix unexpected 4xx and 5xx. Google seulement indexes pages served with a 200, and “Client and server error pages aren’t indexed.”
  • Chaîne de redirectionss and loops. Collapse A→B→C→D bas to A→D. Chains waste explorer and leak a little on every hop. The exploration and redirections material goes deeper ici.
Evidence for this claim A robots.txt disallow controls crawling but is not a reliable mechanism for keeping a linked URL out of Google’s index. Scope: production output verified at URL, template and representative-sample level Confidence: high · Verified: Introduction to robots.txt

2. Indexability

  • Index coverage. In Search Console’s Page Indexation report, reconcile ce que vous vouloir indexé contre ce que en réalité is. Investigate grand “Découvert/Crawled - currently non indexée” buckets.
  • Canonicalization. Point duplicate and near-duplicate URLs at un preferred version. Google calls rel="canonical" “a strong signal que the specified URL devrait become canonical” — a signal, pas a directive it doit obey. And crucially: “Don’t use the robots.txt file for canonicalization purposes.” Vérifier vous aren’t sending conflicting canonical signals à travers HTML tag, HTTP header, and sitemap — the canonicalization deep dive walks via consolidating les.
  • Contenu dupliqué. Parameters, print versions, staging leaks, http/https and www/non-www splits tout créer duplicates. Pick un, canonicalize or redirection the rest.

3. HTTPS

Baseline, pas optional. Serve the whole site over HTTPS, redirection http to https, and hunt bas mixed content (a secure page chargement an insecure image, script, or stylesheet). Chris Green’s SEO in 2026 reality-check pegs HTTPS adoption at “91%+” — vous don’t vouloir to be in the trailing 9%.

4. Mobile SEO

Google uses indexation mobile-first: “Google uses the mobile version of a site’s content, crawled with the smartphone agent, pour indexation and ranking.” The parity checklist, straight from Google’s mobile-first doc:

  • “Make sure that your mobile site contains the same content as your desktop site.”
  • “Assurez-vous que the title element and the meta description are equivalent à travers les deux versions of votre site.”
  • “Make sure that your mobile and desktop sites have the same structured data.”
  • “Use the same robots meta tags on the mobile and desktop site.”
  • “Don’t lazy-load primary content upon user interaction.”
  • “Assurez-vous que the mobile site has the même texte alternatif pour images as the desktop site.”

The la plupart courant échec is a stripped-down mobile template que quietly drops content, liens, or données structurées présent on desktop — Google indexes the thinner version.

5. Sitemaps and discovery

  • XML sitemap hygiene. Utiliser absolute, URL canoniques; stay sous the 50 MB / 50 000-URL per-file limite; liste seulement indexable, URL canoniques; garder lastmod accurate. Référence it in robots.txt (Sitemap: https://example.com/sitemap.xml) so engines découvrir it automatically. Complet treatment in the XML sitemaps material.
  • Bing encore cares. From Bing’s July 2025 guidance: “Sitemaps remain a foundational signal pour ensuring comprehensive URL coverage à travers votre site,” “XML remains the preferred format pour sitemaps,” and “The lastmod field in votre sitemap remains a clé signal, helping Bing prioritize URLs pour recrawling.”
  • IndexNow. La plupart Google-centric checklists omit it, but it’s a live, free, one-line win: Bing’s advice is to “Utiliser IndexNow pour real-time URL submission, instantly notifying Bing and participating moteur de recherches” quand content changements. It complements sitemaps plutôt que replacing les. (Remarque: Google ne fait pas utiliser IndexNow pour general pages.)

6. Données structurées

  • Ce que it fait: earns rich-result eligibility and helps machines (and LLMs) comprendre votre page. “Ajout données structurées peut enable résultats de recherche que are plus engaging to utilisateurs… qui are appelé résultats enrichis.”
  • Ce que it fait pas do: boost rankings. Google’s docs frame schema strictly as rich-result eligibility and machine understanding — pas a ranking signal. Ne faites pas sell it, or budget pour it, as a ranking play.
  • Format: “In general, Google recommends en utilisant JSON-LD pour données structurées si votre site’s setup permet it, as it’s the easiest solution pour website owners to implement and maintain at scale.” Validate with the Résultats enrichis Tester. The données structurées material covers the spécifique types worth implementing.

7. Core Web Vitals and page experience

Targets, pas gates — Google’s réel wording is “strive to,” qui la plupart checklists overstate as hard réussir/échouer cutoffs:

  • LCP“strive to have LCP occur dans the premier 2,5 seconds of lune page starting to charger.”
  • INP“strive to have an INP of less than 200 milliseconds.”
  • CLS“strive to have a CLS score of less than 0.1.”

And the relationship to ranking, qui personnes badly over-weight: “Recherche Google toujours seeks to montrer the la plupart relevant content, même si lune page experience is sub-par.” Bon Core Web Vitals are a tiebreaker among relevant results, pas an override of relevance. Mesurer with real-user (CrUX/field) données, pas simplement lab scores.

8. Avancé additions (grand / ecommerce / JS-heavy seulement)

  • Budget d’exploration. Seulement si vous cleared Google’s thresholds ci-dessus. Capacity + demand; vous gain budget by removing waste (parameter explosions, faceted-nav combinations, spider traps, duplicate URLs) far plus que by trying to faire Google explorer “more.” Voir the budget d’exploration deep dive.
  • JavaScript rendering. Confirmer critical content and liens exist in the rendered HTML and are reachable via réel <a href> liens, pas click-only navigation. Ce is JavaScript SEO territory.
  • Faceted navigation contrôler. Decide qui filter/sort combinations are crawlable/indexable and contrôler the rest.
  • Log-file analysis. The ground truth pour ce que bots en réalité récupérer, how souvent, and ce que code d’états ils hit.
  • hreflang — seulement si you’re genuinely multi-regional/multilingual. That’s a big suffisant topic to live in its propre international-SEO material; don’t bolt it on half-done.

9. AI / LLM robot d’exploration accès (short, scoped)

Two table-stakes items in 2026, and aucun plus — the deep GEO/AEO fonctionner lives in the AI-search material, pas ici:

  • Decide AI-crawler accès in robots.txt. Explicitly autoriser or disallow the AI user-agents vous care à propos de (training vs. AI-search vs. user-triggered fetchers are différent bots). Chris Green’s framing: “Robots.txt is ne … plus simplement explorer housekeeping. It’s becoming a policy surface.”
  • Données structurées doubles as machine context pour LLMs — a bonus raison to obtenir votre schema valid, pas a nouveau workstream.

10. How to prioritize ce que vous trouver

Ce is où la plupart checklists échouer: ils hand vous 90 items with aucun weighting. Don’t fix everything — fix ce que moves the needle. My buddy Patrick’s advice on client audits, qui I garder coming back to: “Si clients are coming to vous asking pour an audit, ils déjà have a pain point. Talk to les. Solve que un chose and they’ll be happy with the audit.” The même SEO Audit Template frames it as “sweating the small stuff rarely does much for your rankings” — meilleur to spend “80% of your time fixing the 20% of things that matter.”

And zoom out on the whole exercise: a checklist obtient vous to okay. Google’s John Mueller has repeatedly made the point que fundamentals alone obtenir vous fine-but-not-great results — réel dominance comes from topical depth and authority, pas from ticking every technical box. The checklist clears the floor; content and liens construire the house.

Vouloir the full-site version?

Ce page is technical-only by design. Si vous vouloir the broader audit — technical plus on-page, content, and off-page — that’s the SEO Audit Checklist, a separate, wider chose. Don’t essayer to faire ce un page do les deux jobs.

Add an expert note

Pin an expert quote

New person? Create their unclaimed profile at /admin/experts/ → Pin a quote first.