SEO teknis Checklist

sebuah comprehensive SEO teknis checklist covering crawlability, pengindeksan, Core Web Vitals, data terstruktur, sitemaps, robots.txt, canonicalization, HTTPS, dan mobile SEO — organized oleh priority dan situs jenis.

Pertama kali diterbitkan: 27 Jun 2026 · Terakhir diperbarui: 3 Agu 2026 · Advanced
Bahasa
1 sinyal bukti di halaman ini

sebuah SEO teknis checklist confirms technical conditions sebuah situs harus meet untuk mesin pencari dan AI jawaban mesin untuk crawl, render, indeks, dan sajikan -nya halaman — Google's four muat-bearing conditions adalah dapat di-crawl, dapat diindeks, understandable, renderable. ini adalah technical-hanya slice; broader full-situs versi lives di SEO Audit Checklist. berfungsi ini di priority order — crawlability, pengindeksan, HTTPS, mobile parity, sitemaps, data terstruktur, Core Web Vitals — dan gate whole bagian oleh situs jenis: sebagian besar kecil situs dapat skip crawl-budget engineering entirely, which Google confirms di -nya own docs. Two myths untuk kill pada sight: robots.txt melakukan not deindex sebuah halaman, dan data terstruktur adalah not sebuah peringkat factor. Meeting checklist adalah necessary tetapi not sufficient — Google says pengindeksan masih isn't guaranteed.

TL;DR — SEO teknis adalah floor, not ceiling: halaman harus menjadi dapat di-crawl → dapat diindeks → understandable → renderable sebelum konten berfungsi dapat pay off. Run checklist di priority order dan gate bagian oleh situs jenis — sebagian besar kecil situs dapat skip anggaran crawling, faceted-nav control, log analysis, dan JS-rendering engineering, which Google’s own crawl-budget doc confirms. Kill two myths pada sight: robots.txt doesn’t deindex, dan data terstruktur isn’t sebuah peringkat factor. Treat Core Web Vitals sebagai targets, not pass/fail gates — itu’s Google’s own “strive to” (terjemahan) “strive untuk” language. dan meeting setiap box masih doesn’t guarantee pengindeksan.

Evidence for this claim Google's minimum technical requirements include accessible Googlebot crawling, a successful HTTP response, and indexable content. Scope: Eligibility prerequisites, not a guarantee of indexing or ranking. Confidence: high · Verified: Google Search Essentials: Technical requirements Evidence for this claim Meeting technical requirements does not guarantee that Google will crawl, index, or serve a page. Scope: Google Search eligibility and selection behavior. Confidence: high · Verified: Google Search Essentials: Technical requirements

mental model: four conditions

Google frames whole dari SEO teknis sekitar four muat-bearing conditions — sebuah halaman memiliki untuk menjadi dapat di-crawl, dapat diindeks, understandable, dan renderable. -nya Search Essentials technical requirements boil itu down untuk three minimums: “Googlebot isn’t blocked,” (terjemahan) “Googlebot isn’t blocked,” “The page works” (terjemahan) “ halaman berfungsi” (disajikan dengan sebuah HTTP 200 status), dan “The page has indexable content.” (terjemahan) “ halaman memiliki dapat diindeks konten.” Everything pada ini checklist adalah really di service dari itu.

critical caveat, straight dari yang sama doc: “Just because a page meets these requirements doesn’t mean that a page will be indexed; indexing isn’t guaranteed.” (terjemahan) “hanya because sebuah halaman meets ini requirements tidak berarti itu sebuah halaman akan menjadi terindeks; pengindeksan isn’t guaranteed.” SEO teknis adalah sebuah gate Anda memiliki untuk pass melalui, not sebuah lever itu guarantees hasil. ini adalah why I’ve selalu argued technical checklist adalah floor — Anda jelas ini so konten dan tautan dapat melakukan mereka job, not alih-alih them.

sebelum Anda start: which bagian bahkan apply untuk Anda?

setiap competing checklist organizes oleh topic dan runs whole thing pada setiap situs. itu’s wrong default. better organizing axis adalah situs jenis, because Google itself gates -nya sebagian besar advanced guidance oleh situs size. dari besar-situs crawl-budget guide: “If your site doesn’t have a large number of pages that change rapidly, or if your pages seem to be crawled the same day that they are published, you don’t need to read this guide.” (terjemahan) “jika Anda situs doesn’t memiliki sebuah besar angka dari halaman itu perubahan rapidly, atau jika Anda halaman seem untuk menjadi di-crawl yang sama day itu mereka adalah published, Anda tidak perlu untuk read ini guide.”

Google’s own (deliberately rough) thresholds untuk when anggaran crawling starts untuk penting:

  • besar situs — 1 million+ unique halaman dengan konten changing tentang weekly.
  • Medium-atau-larger situs — 10 000+ unique halaman dengan very rapidly changing (daily) konten.
  • situs dengan sebuah big share dari URLs stuck di Search Console’s “Discovered - currently not indexed” (terjemahan) “ditemukan - currently not terindeks” state.

So here’s split I’d actually run:

Starter track — kecil / brochure / local-business situs (< ~10K URLs): crawlability sanity periksa, pengindeksan/canonical sanity periksa, HTTPS, mobile parity, one sitemap XML, one round dari Core Web Vitals, valid data terstruktur where ini earns sebuah rich hasil. Skip anggaran crawling, faceted-nav control, log-file analysis, dan JS-rendering engineering entirely.

Advanced track — besar / ecommerce / JS-heavy / enterprise situs: everything di Starter track plus crawl-budget management, faceted-navigation control, JavaScript rendering audits, server-log analysis, dan (jika international) hreflang.

1. Crawlability

  • robots.txt correctness. Confirm Anda aren’t disallowing anything Anda ingin terindeks. Google: “A robots.txt file tells search engine crawlers which URLs the crawler can access on your site. This is used mainly to avoid overloading your site with requests; it is not a mechanism for keeping a web page out of Google.” (terjemahan) “sebuah robots.txt file tells mesin pencari crawler which URLs crawler dapat access pada Anda situs. ini adalah digunakan mainly untuk hindari overloading Anda situs dengan permintaan; ini adalah not sebuah mechanism untuk keeping sebuah halaman web out dari Google.” gunakan ini untuk pertahankan bot out dari rendah-nilai spaces (internal search, infinite parameter combinations), not sebagai sebuah deindexing alat. ini adalah yang sama territory robots.txt deep dive covers di full.
  • ** robots.txt + noindex contradiction trap.** melakukan not block sebuah URL di robots.txt dan rely pada sebuah noindex pada ini. Google: “While Google won’t crawl or index the content blocked by a robots.txt file, we might still find and index a disallowed URL if it is linked from other places on the web.” (terjemahan) “While Google won’t crawl atau indeks konten blocked oleh sebuah robots.txt file, kami mungkin masih temukan dan indeks sebuah disallowed URL jika ini adalah ditautkan dari lainnya places pada web.” bot dapat’t crawl halaman, so ini tidak pernah sees noindex, dan URL dapat masih surface bare di hasil. untuk hapus sebuah halaman: allow crawling + noindex, atau password-protect ini.
  • crawl errors. Fix unexpected 4xx dan 5xx. Google hanya indeks halaman disajikan dengan sebuah 200, dan “Client and server error pages aren’t indexed.” (terjemahan) “Client dan server halaman error aren’t terindeks.”
  • rantai pengalihan dan loops. Collapse sebuah→B→C→D down untuk sebuah→D. Chains waste crawl dan leak sebuah little pada setiap hop. crawling dan redirects material goes deeper here.
Evidence for this claim A robots.txt disallow controls crawling but is not a reliable mechanism for keeping a linked URL out of Google’s index. Scope: production output verified at URL, template and representative-sample level Confidence: high · Verified: Introduction to robots.txt

2. Indexability

  • indeks coverage. di Search Console’s halaman pengindeksan report, reconcile what Anda ingin terindeks terhadap what actually adalah. Investigate besar “Discovered/Crawled - currently not indexed” (terjemahan) “ditemukan/di-crawl - currently not terindeks” buckets.
  • Canonicalization. poin duplicate dan near-duplicate URLs di one preferred versi. Google panggilan rel="canonical" “a strong signal that the specified URL should become canonical” (terjemahan) “sebuah strong signal itu specified URL seharusnya become canonical” — sebuah signal, not sebuah directive ini harus obey. dan crucially: “Don’t use the robots.txt file for canonicalization purposes.” (terjemahan) “Don’t gunakan robots.txt file untuk canonicalization purposes.” periksa Anda aren’t sending conflicting canonical signals di seluruh HTML tag, header HTTP, dan sitemap — canonicalization deep dive walks melalui consolidating them.
  • Duplicate konten. Parameters, print versi, staging leaks, http/https dan www/non-www splits semua buat duplicates. Pick one, canonicalize atau redirect rest.

3. HTTPS

Baseline, not optional. sajikan whole situs di atas HTTPS, redirect http untuk https, dan hunt down mixed konten (sebuah secure halaman memuat sebuah insecure image, script, atau stylesheet). Chris Green’s SEO di 2026 reality-periksa pegs HTTPS adoption di “91%+” (terjemahan) “91%+” — Anda tidak ingin untuk menjadi di trailing 9%.

4. Mobile SEO

Google menggunakan pengindeksan mobile-pertama: “Google uses the mobile version of a site’s content, crawled with the smartphone agent, for indexing and ranking.” (terjemahan) “Google menggunakan mobile versi dari sebuah situs’s konten, di-crawl dengan smartphone agent, untuk pengindeksan dan peringkat.” parity checklist, straight dari Google’s mobile-pertama doc:

  • “Make sure that your mobile site contains the same content as your desktop site.” (terjemahan) “pastikan itu Anda mobile situs berisi yang sama konten sebagai Anda desktop situs.”
  • “Make sure that the title element and the meta description are equivalent across both versions of your site.” (terjemahan) “pastikan itu judul element dan deskripsi meta adalah equivalent di seluruh both versi dari Anda situs.”
  • “Make sure that your mobile and desktop sites have the same structured data.” (terjemahan) “pastikan itu Anda mobile dan desktop situs memiliki yang sama data terstruktur.”
  • “Use the same robots meta tags on the mobile and desktop site.” (terjemahan) “gunakan yang sama robots meta tags pada mobile dan desktop situs.”
  • “Don’t lazy-load primary content upon user interaction.” (terjemahan) “Don’t lazy-muat primary konten upon pengguna interaction.”
  • “Make sure that the mobile site has the same alt text for images as the desktop site.” (terjemahan) “pastikan itu mobile situs memiliki yang sama teks alt untuk images sebagai desktop situs.”

paling umum failure adalah sebuah stripped-down mobile template itu quietly drops konten, tautan, atau data terstruktur present pada desktop — Google indeks thinner versi.

5. Sitemaps dan penemuan

  • sitemap XML hygiene. gunakan absolute, canonical URLs; stay di bawah 50 MB / 50 000-URL per-file limit; list hanya dapat diindeks, canonical URLs; pertahankan lastmod accurate. Reference ini di robots.txt (Sitemap: https://example.com/sitemap.xml) so mesin menemukan ini automatically. Full treatment di sitemap XML material.
  • Bing masih cares. dari Bing’s July 2025 guidance: “Sitemaps remain a foundational signal for ensuring comprehensive URL coverage across your site,” (terjemahan) “Sitemaps remain sebuah foundational signal untuk ensuring comprehensive URL coverage di seluruh Anda situs,” “XML remains the preferred format for sitemaps,” (terjemahan) “XML remains preferred format untuk sitemaps,” dan “The lastmod field in your sitemap remains a key signal, helping Bing prioritize URLs for recrawling.” (terjemahan) “ lastmod field di Anda sitemap remains sebuah key signal, helping Bing prioritize URLs untuk recrawling.”
  • IndexNow. sebagian besar Google-centric checklists omit ini, tetapi ini adalah sebuah live, free, one-line win: Bing’s advice adalah untuk “Use IndexNow for real-time URL submission, instantly notifying Bing and participating search engines” (terjemahan) “gunakan IndexNow untuk nyata-time URL submission, instantly notifying Bing dan participating mesin pencari” when konten perubahan. ini complements sitemaps alih-alih replacing them. (Note: Google melakukan not gunakan IndexNow untuk umum halaman.)

6. data terstruktur

  • What ini melakukan: earns rich-hasil eligibility dan helps machines (dan LLMs) memahami Anda halaman. “Adding structured data can enable search results that are more engaging to users… which are called rich results.” (terjemahan) “menambahkan data terstruktur dapat enable hasil pencarian itu adalah more engaging untuk pengguna… which adalah called rich hasil.”
  • What ini melakukan not melakukan: boost rankings. Google’s docs frame schema strictly sebagai rich-hasil eligibility dan machine understanding — not sebuah sinyal peringkat. melakukan not sell ini, atau budget untuk ini, sebagai sebuah peringkat play.
  • Format: “In general, Google recommends using JSON-LD for structured data if your site’s setup allows it, as it’s the easiest solution for website owners to implement and maintain at scale.” (terjemahan) “di umum, Google recommends menggunakan JSON-LD untuk data terstruktur jika Anda situs’s setup allows ini, sebagai ini adalah easiest solusi untuk situs web owners untuk implement dan maintain di scale.” Validate dengan Rich hasil Test. data terstruktur material covers spesifik jenis worth implementing.

7. Core Web Vitals dan pengalaman halaman

Targets, not gates — Google’s actual wording adalah “strive to,” (terjemahan) “strive untuk,” which sebagian besar checklists overstate sebagai hard pass/fail cutoffs:

  • LCP“strive to have LCP occur within the first 2.5 seconds of the page starting to load.” (terjemahan) “strive untuk memiliki LCP occur di dalam pertama 2,5 seconds dari halaman starting untuk muat.”
  • INP“strive to have an INP of less than 200 milliseconds.” (terjemahan) “strive untuk memiliki sebuah INP dari less daripada 200 milliseconds.”
  • CLS“strive to have a CLS score of less than 0.1.” (terjemahan) “strive untuk memiliki sebuah CLS score dari less daripada 0,1.”

dan relationship untuk peringkat, which people badly di atas-weight: “Google Search always seeks to show the most relevant content, even if the page experience is sub-par.” (terjemahan) “Google Search selalu seeks untuk tampilkan paling relevant konten, bahkan jika pengalaman halaman adalah sub-par.” baik Core Web Vitals adalah sebuah tiebreaker among relevant hasil, not sebuah override dari relevance. mengukur dengan nyata-pengguna (CrUX/field) data, not hanya lab scores.

8. Advanced additions (besar / ecommerce / JS-heavy hanya)

  • anggaran crawling. hanya jika Anda cleared Google’s thresholds above. Capacity + demand; Anda gain budget oleh menghapus waste (parameter explosions, faceted-nav combinations, spider traps, duplicate URLs) far more daripada oleh trying untuk membuat Google crawl “more.” (terjemahan) “more.” See anggaran crawling deep dive.
  • JavaScript rendering. Confirm critical konten dan tautan exist di rendered HTML dan adalah reachable via nyata <a href> tautan, not click-hanya navigation. ini adalah JavaScript SEO territory.
  • Faceted navigation control. Decide which filter/sort combinations adalah dapat di-crawl/dapat diindeks dan control rest.
  • Log-file analysis. ground truth untuk what bot actually fetch, how sering, dan what kode status mereka hit.
  • hreflang — hanya jika Anda’re genuinely multi-regional/multilingual. itu’s sebuah big enough topic untuk live di -nya own international-SEO material; don’t bolt ini pada half-done.

9. AI / LLM crawler access (pendek, scoped)

Two table-stakes items di 2026, dan no more — deep GEO/AEO berfungsi lives di AI-search material, not here:

  • Decide AI-crawler access di robots.txt. Explicitly allow atau disallow AI pengguna-agents Anda care tentang (training vs. AI-search vs. pengguna-triggered fetchers adalah berbeda bot). Chris Green’s framing: “Robots.txt is no longer just crawl housekeeping. It’s becoming a policy surface.” (terjemahan) “Robots.txt adalah no longer hanya crawl housekeeping. ini adalah becoming sebuah policy surface.”
  • data terstruktur doubles sebagai machine context untuk LLMs — sebuah bonus alasan untuk get Anda schema valid, not sebuah baru workstream.

10. cara prioritize what Anda temukan

ini adalah where sebagian besar checklists fail: mereka hand Anda 90 items dengan no weighting. Don’t fix everything — fix what moves needle. My buddy Patrick’s advice pada client audits, which I pertahankan coming back untuk: “If clients are coming to you asking for an audit, they already have a pain point. Talk to them. Solve that one thing and they’ll be happy with the audit.” (terjemahan) “jika clients adalah coming untuk Anda asking untuk sebuah audit, mereka sudah memiliki sebuah pain poin. Talk untuk them. Solve itu one thing dan mereka’ll menjadi happy dengan audit.” yang sama SEO Audit Template frames ini sebagai “sweating the small stuff rarely does much for your rankings” (terjemahan) “sweating kecil stuff rarely melakukan much untuk Anda rankings” — better untuk spend “80% of your time fixing the 20% of things that matter.” (terjemahan) “80% dari Anda time fixing 20% dari things itu penting.”

dan zoom out pada whole exercise: sebuah checklist gets Anda untuk okay. Google’s John Mueller memiliki repeatedly dibuat poin itu fundamentals alone get Anda fine-tetapi-not-great hasil — nyata dominance comes dari topical depth dan authority, not dari ticking setiap technical box. checklist clears floor; konten dan tautan bangun house.

ingin full-situs versi?

ini halaman adalah technical-hanya oleh design. jika Anda ingin broader audit — technical plus pada-halaman, konten, dan off-halaman — itu’s SEO Audit Checklist, sebuah separate, wider thing. Don’t try untuk membuat ini one halaman melakukan both jobs.

Add an expert note

Pin an expert quote

New person? Create their unclaimed profile at /admin/experts/ → Pin a quote first.