SEO kỹ thuật Checklist

MỘT comprehensive SEO kỹ thuật checklist covering crawlability, lập chỉ mục, Core Web Vitals, dữ liệu có cấu trúc, sitemaps, robots.txt, canonicalization, HTTPS, và mobile SEO — organized by priority và site loại.

Xuất bản lần đầu: 27 thg 6, 2026 · Cập nhật lần cuối: 8 thg 8, 2026 · Advanced
Ngôn ngữ
1 tín hiệu bằng chứng trên trang này

MỘT SEO kỹ thuật checklist xác nhận đó kỹ thuật conditions một site phải đáp ứng cho các công cụ tìm kiếm và AI câu trả lời engines để crawl, render, chỉ mục, và serve của nó các trang — Google four load-bearing conditions là crawlable, indexable, understandable, renderable. đây là đó kỹ thuật-chỉ slice; đó rộng hơn đầy đủ-site version lives trong đó SEO Audit Checklist. Hoạt động điều này trong priority order — crawlability, lập chỉ mục, HTTPS, mobile parity, sitemaps, dữ liệu có cấu trúc, Core Web Vitals — và gate toàn bộ sections by site loại: hầu hết nhỏ các trang có thể skip crawl-budget engineering hoàn toàn, mà Google xác nhận trong của nó own tài liệu. Hai myths để kill on sight: robots.txt không deindex một trang, và dữ liệu có cấu trúc không phải một xếp hạng factor. Đáp ứng đó checklist là necessary nhưng không sufficient — Google says lập chỉ mục vẫn không guaranteed.

Tóm tắt — kỹ thuật SEO là floor, không ceiling: các trang phải là crawlable → indexable → understandable → renderable trước khi nội dung hoạt động có thể pay off. Chạy checklist trong priority order và gate sections by trang web loại — phần lớn nhỏ các trang có thể skip ngân sách crawl, faceted-nav control, log analysis, và JS-kết xuất engineering, mà Google own crawl-budget doc xác nhận. Kill hai myths on sight: robots.txt không deindex, và dữ liệu có cấu trúc không phải xếp hạng factor. Treat Core Web Vitals as targets, không truyền/fail gates — đó Google own “strive để” language. và đáp ứng mỗi box vẫn không bảo đảm lập chỉ mục.

Evidence for this claim Google's minimum technical requirements include accessible Googlebot crawling, a successful HTTP response, and indexable content. Scope: Eligibility prerequisites, not a guarantee of indexing or ranking. Confidence: high · Verified: Google Search Essentials: Technical requirements Evidence for this claim Meeting technical requirements does not guarantee that Google will crawl, index, or serve a page. Scope: Google Search eligibility and selection behavior. Confidence: high · Verified: Google Search Essentials: Technical requirements

mental model: four conditions

Google frames đó toàn bộ of SEO kỹ thuật khoảng four load-bearing conditions — một trang có để là crawlable, indexable, understandable, và renderable. Của nó Tìm kiếm Essentials kỹ thuật requirements boil đó xuống để three minimums: “Googlebot isn’t blocked,” (bản dịch) «Googlebot không blocked,» “The page works” (bản dịch) «Đó trang hoạt động» (phân phối với an HTTP 200 status), và “The page has indexable content.” (bản dịch) «Đó trang có indexable nội dung.» Mọi thứ on này checklist là thực sự trong service of những.

Đó cốt yếu caveat, straight từ đó giống nhau doc: “Just because a page meets these requirements doesn’t mean that a page will be indexed; indexing isn’t guaranteed.” (bản dịch) «Chỉ vì một trang đáp ứng những requirements không có nghĩa là đó một trang sẽ là được lập chỉ mục; lập chỉ mục không guaranteed.» SEO kỹ thuật là một gate bạn có để truyền qua, không một lever đó bảo đảm kết quả. Này là vì sao I’ve luôn argued đó kỹ thuật checklist là đó floor — bạn clear điều này so nội dung và links có thể làm của họ job, không thay vì them.

trước khi bạn bắt đầu: mà sections ngay cả apply để bạn?

Mỗi competing checklist organizes by topic và chạy đó toàn bộ điều on mỗi site. đó là đó sai default. Đó tốt hơn organizing axis là site loại, vì Google itself gates của nó hầu hết advanced hướng dẫn by site size. Từ đó lớn-site crawl-budget hướng dẫn: “If your site doesn’t have a large number of pages that change rapidly, or if your pages seem to be crawled the same day that they are published, you don’t need to read this guide.” (bản dịch) «Nếu trang web của bạn không có nhiều trang thay đổi nhanh, hoặc nếu các trang của bạn dường như được crawl ngay trong ngày xuất bản, bạn không cần đọc hướng dẫn này.»

Google own (có chủ ý rough) thresholds cho Khi ngân sách crawl bắt đầu để quan trọng:

  • Lớn các trang — 1 million+ unique các trang với nội dung thay đổi về weekly.
  • Medium-hoặc-lớn hơn các trang — 10 000+ unique các trang với very rapidly thay đổi (daily) nội dung.
  • Các trang với một big share of URLs stuck trong Search Console’s “Discovered - currently not indexed” (bản dịch) «Discovered - hiện tại không được lập chỉ mục» state.

So ở đây split I’d thực ra chạy:

Starter track — nhỏ / brochure / local-business các trang (< ~10K các URL): crawlability sanity kiểm tra, lập chỉ mục/canonical sanity kiểm tra, HTTPS, mobile parity, một XML sitemap, một round của Core Web Vitals, hợp lệ dữ liệu có cấu trúc nơi nó earns rich kết quả. Skip ngân sách crawl, faceted-nav control, log-file analysis, và JS-kết xuất engineering hoàn toàn.

Nâng cao track — lớn / ecommerce / JS-nặng / enterprise các trang: mọi thứ trong Starter track plus crawl-budget management, faceted-navigation control, JavaScript kết xuất audits, máy chủ-log analysis, và (nếu international) hreflang.

1. Crawlability

  • robots.txt correctness. Xác nhận bạn không disallowing bất cứ điều gì bạn muốn được lập chỉ mục. Google: “A robots.txt file tells search engine crawlers which URLs the crawler can access on your site. This is used mainly to avoid overloading your site with requests; it is not a mechanism for keeping a web page out of Google.” (bản dịch) «MỘT robots.txt file tells công cụ tìm kiếm các crawler mà URLs đó crawler có thể access trên trang web của bạn. Này là dùng mainly để tránh overloading trang web của bạn với các yêu cầu; điều này không phải một mechanism cho giữ một web trang out of Google.» Dùng điều này để giữ bots out of thấp-giá trị spaces (internal tìm kiếm, infinite parameter combinations), không as một deindexing tool. Này là đó giống nhau territory đó robots.txt deep dive covers trong đầy đủ.
  • Đó robots.txt + noindex contradiction trap. Làm không block một URL trong robots.txt rely on một noindex on điều này. Google: “While Google won’t crawl or index the content blocked by a robots.txt file, we might still find and index a disallowed URL if it is linked from other places on the web.” (bản dịch) «Trong khi Google sẽ không crawl hoặc chỉ mục đó nội dung blocked by một robots.txt file, we có thể vẫn tìm và chỉ mục một disallowed URL nếu điều này là linked từ other places on đó web.» Đó bot không thể crawl đó trang, so điều này không bao giờ sees đó noindex, và đó URL có thể vẫn surface bare trong kết quả. Để xóa một trang: cho phép crawling + noindex, hoặc password-bảo vệ điều này.
  • Crawl các lỗi. Cách sửa unexpected 4xx và 5xx. Google chỉ indexes các trang phân phối với một 200, và “Client and server error pages aren’t indexed.” (bản dịch) «Client và máy chủ lỗi các trang không được lập chỉ mục.»
  • Chuyển hướng chains và loops. Collapse MỘT→B→C→D xuống để MỘT→D. Chains waste crawl và leak một little on mỗi hop. Đó crawling và các chuyển hướng material goes deeper ở đây.
Evidence for this claim A robots.txt disallow controls crawling but is not a reliable mechanism for keeping a linked URL out of Google’s index. Scope: production output verified at URL, template and representative-sample level Confidence: high · Verified: Introduction to robots.txt

2. Indexability

  • Chỉ mục coverage. Trong Search Console’s Trang Lập chỉ mục báo cáo, reconcile điều gì bạn muốn được lập chỉ mục so với điều gì thực ra là. Investigate lớn “Discovered/Crawled - currently not indexed” (bản dịch) «Discovered/Được crawl - hiện tại không được lập chỉ mục» buckets.
  • Canonicalization. Point duplicate và near-duplicate URLs tại một được ưu tiên version. Google calls rel="canonical" “a strong signal that the specified URL should become canonical” (bản dịch) «một mạnh tín hiệu đó specified URL nên become canonical» — một tín hiệu, không một directive điều này phải obey. Và crucially: “Don’t use the robots.txt file for canonicalization purposes.” (bản dịch) «không dùng đó robots.txt file cho canonicalization purposes.» Kiểm tra bạn không sending conflicting các tín hiệu canonical trên HTML tag, HTTP header, và sitemap — đó canonicalization deep dive walks qua consolidating them.
  • Duplicate nội dung. Parameters, print versions, staging leaks, http/httpswww/non-www splits all tạo duplicates. Pick một, canonicalize hoặc chuyển hướng đó rest.

3. HTTPS

Baseline, không tùy chọn. phục vụ toàn bộ trang web over HTTPS, chuyển hướng http để https, và hunt xuống mixed nội dung ( secure trang loading insecure image, script, hoặc stylesheet). Chris Green SEO trong 2026 reality-kiểm tra pegs HTTPS adoption tại “91%+” — bạn không muốn để là trong trailing 9%.

4. Mobile SEO

Google dùng mobile-đầu tiên lập chỉ mục: “Google uses the mobile version of a site’s content, crawled with the smartphone agent, for indexing and ranking.” (bản dịch) «Google dùng đó mobile version of một site nội dung, được crawl với đó smartphone agent, cho lập chỉ mục và xếp hạng.» Đó parity checklist, straight từ Google mobile-đầu tiên doc:

  • “Make sure that your mobile site contains the same content as your desktop site.” (bản dịch) «Hãy bảo đảm đó của bạn mobile site contains đó giống nhau nội dung as của bạn desktop site.»
  • “Make sure that the title element and the meta description are equivalent across both versions of your site.” (bản dịch) «Hãy bảo đảm đó tiêu đề element và đó mô tả meta là tương đương trên cả hai versions of trang web của bạn.»
  • “Make sure that your mobile and desktop sites have the same structured data.” (bản dịch) «Hãy bảo đảm đó của bạn mobile và desktop các trang có đó giống nhau dữ liệu có cấu trúc.»
  • “Use the same robots meta tags on the mobile and desktop site.” (bản dịch) «Dùng đó giống nhau robots meta tags on đó mobile và desktop site.»
  • “Don’t lazy-load primary content upon user interaction.” (bản dịch) «không lazy-load chính nội dung upon người dùng interaction.»
  • “Make sure that the mobile site has the same alt text for images as the desktop site.” (bản dịch) «Hãy bảo đảm đó mobile site có đó giống nhau alt text cho images as đó desktop site.»

phần lớn phổ biến thất bại là stripped-xuống mobile template đó âm thầm drops nội dung, links, hoặc dữ liệu có cấu trúc present on desktop — Google indexes thinner version.

5. Sitemaps và phát hiện

  • XML sitemap hygiene. Dùng absolute, canonical URLs; stay dưới đó 50 MB / 50 000-URL theo-file limit; list chỉ indexable, canonical URLs; giữ lastmod chính xác. Reference điều này trong robots.txt (Sitemap: https://example.com/sitemap.xml) so engines discover điều này tự động. Đầy đủ treatment trong đó XML sitemaps material.
  • Bing vẫn cares. Từ Bing July 2025 hướng dẫn: “Sitemaps remain a foundational signal for ensuring comprehensive URL coverage across your site,” (bản dịch) «Sitemaps vẫn một foundational tín hiệu cho bảo đảm comprehensive URL coverage trên trang web của bạn,» “XML remains the preferred format for sitemaps,” (bản dịch) «XML vẫn đó được ưu tiên format cho sitemaps,»“The lastmod field in your sitemap remains a key signal, helping Bing prioritize URLs for recrawling.” (bản dịch) «Đó lastmod trường trong của bạn sitemap vẫn một key tín hiệu, helping Bing prioritize URLs cho recrawling.»
  • IndexNow. Hầu hết Google-centric checklists omit điều này, nhưng đây là một trực tiếp, free, một-line win: Bing advice là để “Use IndexNow for real-time URL submission, instantly notifying Bing and participating search engines” (bản dịch) «Dùng IndexNow cho real-time URL submission, instantly notifying Bing và participating các công cụ tìm kiếm» khi nội dung thay đổi. Điều này complements sitemaps thay vì thay thế them. (Note: Google không dùng IndexNow cho chung các trang.)

6. Dữ liệu có cấu trúc

  • Điều gì điều này làm: earns rich-kết quả eligibility và helps machines (và LLMs) understand trang của bạn. “Adding structured data can enable search results that are more engaging to users… which are called rich results.” (bản dịch) «Thêm dữ liệu có cấu trúc có thể enable kết quả tìm kiếm đó là hơn engaging để người dùng… mà là called rich kết quả.»
  • Điều gì điều này làm không làm: boost thứ hạng. Google tài liệu frame schema strictly as rich-kết quả eligibility và machine understanding — không phải là tín hiệu xếp hạng. Không sell điều này, hoặc budget cho điều này, as một xếp hạng play.
  • Format: “In general, Google recommends using JSON-LD for structured data if your site’s setup allows it, as it’s the easiest solution for website owners to implement and maintain at scale.” (bản dịch) «Nhìn chung, Google khuyến nghị dùng JSON-LD cho dữ liệu có cấu trúc nếu trang web của bạn setup cho phép điều này, as đây là đó easiest giải pháp cho website owners để implement và maintain tại quy mô.» Validate với đó Rich Kết quả Kiểm thử. Đó dữ liệu có cấu trúc material covers đó cụ thể types worth implementing.

7. Core Web Vitals và trang experience

Targets, không gates — Google thực tế wording là “strive để,” mà phần lớn checklists overstate as hard truyền/fail cutoffs:

  • LCP“strive to have LCP occur within the first 2.5 seconds of the page starting to load.” (bản dịch) «strive để có LCP occur trong đó đầu tiên 2,5 seconds of đó trang starting để load.»
  • INP“strive to have an INP of less than 200 milliseconds.” (bản dịch) «strive để có an INP of ít hơn 200 milliseconds.»
  • CLS“strive to have a CLS score of less than 0.1.” (bản dịch) «strive để có một CLS score of ít hơn 0,1.»

Và đó mối quan hệ để xếp hạng, mà mọi người badly over-weight: “Google Search always seeks to show the most relevant content, even if the page experience is sub-par.” (bản dịch) «Google Search luôn seeks để cho thấy đó hầu hết relevant nội dung, ngay cả khi đó trang experience là sub-par.» Good Core Web Vitals là một tiebreaker among relevant kết quả, không an override of relevance. Đo lường với real-người dùng (CrUX/trường) dữ liệu, không chỉ lab scores.

8. Nâng cao additions (lớn / ecommerce / JS-nặng chỉ)

  • Ngân sách crawl. chỉ nếu bạn cleared Google thresholds trên. Capacity + demand; bạn gain budget by removing waste (parameter explosions, faceted-nav combinations, spider traps, duplicate các URL) far nhiều hơn by trying để làm Google crawl “hơn.” See ngân sách crawl deep dive.
  • JavaScript kết xuất. xác nhận cốt yếu nội dung và links exist trong được kết xuất HTML và là reachable qua thực <a href> links, không nhấp-chỉ navigation. Đây là JavaScript SEO territory.
  • Faceted navigation control. quyết định mà filter/loại combinations là crawlable/indexable và control rest.
  • Log-file analysis. ground truth cho Điều gì bots thực ra fetch, Cách thường, và Điều gì các mã trạng thái họ hit.
  • hreflang — chỉ nếu bạn’re genuinely multi-regional/multilingual. đó big đủ topic để trực tiếp trong của nó own international-SEO material; không bolt nó on half-đã xong.

9. AI / LLM crawler access (ngắn, scoped)

Hai bảng-stakes items trong 2026, và không nhiều hơn — deep GEO/AEO hoạt động lives trong AI-tìm kiếm material, không ở đây:

  • Decide AI-crawler access trong robots.txt. Explicitly cho phép hoặc disallow đó AI người dùng-agents bạn care về (training so với. AI-tìm kiếm so với. người dùng-triggered fetchers là khác nhau bots). Chris Green cách diễn đạt: “Robots.txt is no longer just crawl housekeeping. It’s becoming a policy surface.” (bản dịch) «Robots.txt là không lâu hơn chỉ crawl housekeeping. đây là becoming một policy surface.»
  • Dữ liệu có cấu trúc doubles as machine context cho LLMs — một bonus reason để nhận của bạn schema hợp lệ, không một new workstream.

10. Cách prioritize Điều gì bạn tìm

Này là nơi hầu hết checklists fail: they hand bạn 90 items với không weighting. không cách sửa mọi thứ — cách sửa điều gì moves đó needle. My buddy Patrick advice on client audits, mà I giữ coming lại để: “If clients are coming to you asking for an audit, they already have a pain point. Talk to them. Solve that one thing and they’ll be happy with the audit.” (bản dịch) «Nếu clients là coming để bạn asking cho an audit, they đã có một pain point. Talk để them. Solve đó một điều và they’ll là happy với đó audit.» Đó giống nhau SEO Audit Template frames điều này as “sweating the small stuff rarely does much for your rankings” (bản dịch) «sweating đó nhỏ stuff rarely làm nhiều cho của bạn thứ hạng» — tốt hơn để spend “80% of your time fixing the 20% of things that matter.” (bản dịch) «80% of của bạn time sửa đó 20% of điều đó quan trọng.»

và zoom out on toàn bộ exercise: checklist nhận bạn để okay. Google John Mueller có repeatedly đã làm point đó fundamentals alone nhận bạn fine-nhưng-không-great kết quả — thực dominance xuất hiện từ topical depth và authority, không từ ticking mỗi kỹ thuật box. checklist clears floor; nội dung và links xây dựng house.

Muốn đầy đủ-trang web version?

điều này trang là kỹ thuật-chỉ by design. nếu bạn muốn rộng hơn audit — kỹ thuật plus on-trang, nội dung, và off-trang — đó SEO Audit Checklist, tách biệt, wider điều. không try để làm điều này một trang làm cả hai jobs.

Add an expert note

Pin an expert quote

New person? Create their unclaimed profile at /admin/experts/ → Pin a quote first.