Thin Content

What thin content really is (value per page, not word count), the difference between the manual action and algorithmic suppression, and why it's a programmatic SEO risk.

First published: Jun 26, 2026 · Last updated: Jul 18, 2026 · Advanced
demand #2 in Risks#3 in Programmatic SEO#207 on the site

Thin content is content that adds little or no value — and it's about value per page, not word count. A focused 200-word answer can be great; a padded 3,000-word article can be thin. There are two separate enforcement paths people constantly conflate: the 'thin content with little or no added value' manual action (a human reviewer, lifted via reconsideration request) and algorithmic suppression (Panda's legacy now in core ranking, plus the absorbed Helpful Content System — no notification, no reconsideration). Google names four subtypes: scaled content abuse, scraping, thin affiliation, and doorways. The dangerous part for programmatic SEO is that thin content is usually a site-level problem, so one cluster of thin templated pages can drag down the whole domain. My fix is always the same: improve with original value, consolidate, or remove — adding words doesn't fix a value problem.

TL;DR — Thin contentThin content is web content that provides little or no value to users. Google's spam policies name it 'thin content with little or no added value' — and it's about value per page, not word count. is content that adds little or no value, defined by value per page — not word count. There are two enforcement paths people conflate: the “thin content with little or no added value” manual action (a human reviewer; lifted by a reconsideration request) and algorithmic suppression (Panda’s legacy, now in core ranking, plus the Helpful Content SystemThe Helpful Content Update (HCU) was a series of Google updates starting in August 2022 that added a site-wide, machine-learning classifier to demote content made primarily to rank rather than to help people. In March 2024 it was folded into Google's core ranking system. Google folded into the March 2024 core update — no notification, no reconsideration). Google names four subtypes: scaled content abuseScaled content abuse is Google's spam policy (introduced March 2024) for generating many low-value pages primarily to manipulate search rankings rather than help users — and it applies no matter how the content is created: AI, automation, or human writers., scraping, thin affiliation, doorways. The dangerous part for programmatic SEOProgrammatic SEO (pSEO) is the practice of generating many pages from a single template plus a data source to target large sets of similar queries. It's powerful when each page genuinely answers its query with unique data, and spam when it just stamps a thin template across a shallow dataset. is that thin content is usually a site-level problem — one cluster of thin templated pages can drag the whole domain down. The fix is always improve, consolidate, or remove; adding words fixes nothing.

What “thin content” actually means

Thin content is not synonymous with short content; the documented issue is lack of added value. Evidence for this claim Official or primary documentation supporting the adjacent article claim, with scope limited to the source's published description. Scope: No ranking guarantee or undisclosed system mechanics are inferred beyond the cited source. Confidence: high · Verified: Google: Thin content manual action Ranking and indexingStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed. outcomes are contextual, so no single page-length threshold guarantees success or failure. Evidence for this claim Official or primary documentation supporting the adjacent article claim, with scope limited to the source's published description. Scope: No ranking guarantee or undisclosed system mechanics are inferred beyond the cited source. Confidence: high · Verified: Google: Helpful content guidance

The mistake I see most: people equate “thin” with “short” and go off to pad their pages. Wrong target. Thin content is about value, not volume. A 150-word page that definitively answers a narrow question is not thin. A 2,000-word article that restates obvious information is. John Mueller has said repeatedly that word count is not a ranking factor, and that holds here — Google evaluates informational value, not length.

So what does “added value” actually mean? In practice, for Google it’s: original information, first-hand experience, original analysis or research, proprietary data, a unique tool or feature, or genuine editorial judgment — something a reader can’t easily get in the same form elsewhere. Simply reorganizing existing information does not qualify. That’s the test, and it’s worth applying brutally to your own pages.

The four official subtypes

Google’s spam policies name specific thin-content types, and they have different remedies:

Scaled content abuse

Google’s words: Scaled content abuseScaled content abuse is Google's spam policy (introduced March 2024) for generating many low-value pages primarily to manipulate search rankings rather than help users — and it applies no matter how the content is created: AI, automation, or human writers. is when many pages are generated for the primary purpose of manipulating search rankings and not helping users.” This is the one that matters most for the modern web. In the March 2024 spam policy update Google renamed “spammy auto-generated content” to “scaled content abuse” — explicitly to cover AI-generated pages at scale. The examples Google lists include using generative AI to “generate many pages without adding value,” scraping feeds to generate pages through “synonymizing, translating, or other obfuscation techniques,” and “creating many pages where the content makes little or no sense to a reader but contains search keywords.”

Scraping

Taking content from other sites and hosting it to manipulate rankings — republishing without original additions, slightly modifying via synonyms or automated rewriting, or compiling other people’s media “without substantial added value.”

Thin affiliation

Google calls it out by name: affiliate pages “where the product descriptions and reviews are copied directly from the original merchant without any original content or added value.” The key nuance: the issue is not having affiliate links. “Good affiliate sites add value by offering meaningful content or features.” The problem is having only affiliate content with nothing of your own.

Doorway pages

Templated pages built mainly to funnel users to one destination — the classic [service] in [city] funnel where every page is substantially the same.

How Google enforces it: two separate mechanisms

This is the distinction almost every other article on thin content gets wrong, so I’m going to be explicit.

Algorithmic suppression (the common one)

Google launched Panda in February 2011 specifically to algorithmically demote sites with thin or low-quality content — content farms and affiliate sites with no editorial value. Through 2011 and 2012 it ran as near-monthly refreshes; by 2016 it was fully folded into Google’s core ranking signals and stopped being named.

Then in August 2022 came the Helpful Content UpdateThe Helpful Content Update (HCU) was a series of Google updates starting in August 2022 that added a site-wide, machine-learning classifier to demote content made primarily to rank rather than to help people. In March 2024 it was folded into Google's core ranking system., a site-wide ML classifier that judged the overall content health of a domain, not just individual pages. In the March 2024 core update, Google deprecated the Helpful Content System as a standalone signal and folded it into core ranking, stating the integration would reduce low-quality, unhelpful content in results by 40 percent. Sites with high proportions of thin, templated, or low-value AI pages were among the hardest hit.

The defining feature of algorithmic suppression: no notification, no reconsideration request. Your pages just stop ranking. You recover by genuinely improving quality and waiting for the next crawl or core update.

The manual action (the loud one)

“Thin content with little or no added value” is a site-level or partial-site manual action applied by a human on Google’s Search quality team — not an algorithm. It appears in the Manual Actions report in Search ConsoleA free Google service that reports how a site performs in Google Search and surfaces problems with how Google crawls, indexes, and serves it. It's first-party data straight from Google — but you don't need it to appear in results.. Google describes the target as content that “doesn’t provide users with substantially unique or valuable content” in violation of the spam policies. Common targets: thin affiliate pages, scraped content, doorways, and templated pages with minimal distinguishing content.

Recovery here is a defined process: identify every affected page, either substantively improve them or remove/noindexNoindex is a directive that tells search engines to keep a page out of their index, so it won't appear in search results. It works only on pages a crawler can actually fetch — a page blocked in robots.txt can never be noindexed. and consolidate, then submit a reconsideration request in Search ConsoleGoogle's free tool for monitoring crawling, indexing, and search performance..

Why the distinction matters: a manual action needs a reconsideration request to lift; algorithmic suppression does not (and won’t respond to one). If you don’t have a manual action in Search Console, you are not automatically “fine” — algorithmic suppression is the far more common outcome and it sends no notice.

Why thin content is the programmatic SEO risk

Here’s the bit that earns this page its place in the Risks cluster. Thin content issues are usually site-level, not page-level. As Mueller put it, “it is not so much that one page doesn’t have enough content, but instead, the website overall is very light in content.”

That site-level nature is exactly what makes programmatic SEOProgrammatic SEO (pSEO) is the practice of generating many pages from a single template plus a data source to target large sets of similar queries. It's powerful when each page genuinely answers its query with unique data, and spam when it just stamps a thin template across a shallow dataset. dangerous. When you generate thousands of pages from one template, you’re not risking one page — you’re risking the whole domain’s quality score. A cluster of thin templated pages can drag down everything, including the good pages you wrote by hand. That’s why I treat thin content as the central risk of any programmatic project, not a footnote.

Mueller’s blunt version: “Programmatic SEO is often a fancy banner for spam.” He’s careful to add it isn’t always spam — and that’s the whole point. The technique is neutral. The pSEO success stories (Wise’s currency pages, Zapier’s integration pages, Nomadlist’s city pages) work because each page carries proprietary, genuinely unique data. The failures happen when the only unique element on a page is the keyword.

My one-line test, the same one I use across the programmatic-SEO pillar: take a page and mentally delete the modifier you’re targeting. If what’s left is a generic, could-be-anything page, it’s thin. And the fix is a data problem, not a word-count problem — add depth and dimensions only you have, or don’t publish that page.

And to be clear, because people get this backwards: AI content is not automatically thin. Google’s policy targets AI content generated “without adding value” — not all AI content. The test is identical to everything else here: does the page give the reader something useful they can’t easily get elsewhere?

How to find thin content on your site

  • Search Console first. Check the Manual Actions report for an actual penalty with a named scope. Without an entry there, don’t call a ranking or traffic decline a “thin content penalty” — a manual action is a specific, reported enforcement state, not a label for any drop.
  • Rule out other explanations before blaming algorithmic suppression. A traffic drop that lines up with a core or HCU update is a reasonable signal, but check first that it isn’t a canonical-selection change (Google picked a different URL as representative), a crawl or indexingStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed. problem, or a genuine demand change — those need different fixes than a content-value problem.
  • Crawl for the signals. Use a crawlerA crawler — also called a spider or bot — is an automated program that fetches web pages, extracts their links, and queues new URLs to visit. Search engines use crawlers to discover and download content for their index. (Ahrefs Site Audit, Screaming Frog) to surface low word-count pages, high cross-page duplication, and a low ratio of unique-to-template content. Word count alone isn’t the verdict — it’s a flag to go look.
  • Run a task-satisfaction check per page, not just an aggregate score. For a sample of pages, ask: what task or question is this page for, does it fully answer it, what evidence or unique contribution does it offer, and what’s the reader’s likely next question? Traffic or word count alone can’t answer that — it takes reading the page against its intended task.
  • Watch the indexing reports.Crawled – currently not indexedA Google Search Console Page Indexing status meaning Googlebot fetched the page but Google decided not to index it — usually a content- or site-quality signal, not a technical error.” and “Discovered – currently not indexed” creeping up across a templated set is Google telling you those pages aren’t worth its space. That’s usually a data-depth problem.

How to fix it

Three moves, in order of preference per page:

  1. Improve — add original data, first-hand analysis, real experience, or a genuine tool or feature. Adding words without adding information does nothing; Google evaluates value, not length.
  2. Consolidate — merge several overlapping thin pages into one comprehensive page that actually deserves to rank.
  3. Remove / noindexNoindex is a directive that tells search engines to keep a page out of their index, so it won't appear in search results. It works only on pages a crawler can actually fetch — a page blocked in robots.txt can never be noindexed. — for pages with no redeeming value and no meaningful traffic, take them out. Before you do, check for dependencies: internal linksAn internal link is a hyperlink from one page on a website to another page on the same website. Internal links help search engines discover your pages and pass ranking signals (PageRank and anchor-text context) between them. pointing to the page, backlinks worth redirecting rather than losing, and any residual traffic or conversions the page still serves. Pruning dead weight protects the rest of the domain, but none of these three moves is a guaranteed fix — verify the outcome after the change rather than assuming it worked.

If — and only if — you have a manual action, finish with a reconsideration request once the cleanup is done.

A few myths to kill

  • “Thin means short.” No. Depth of value is the metric.
  • “Only AI content is thin.” No — scraped, spun, and affiliate-duplicated content have been targeted since 2011, long before modern AI.
  • “No manual action means I’m fine.” No — algorithmic suppression is silent and far more common.
  • “More words fixes it.” No — words aren’t value.
  • “Disavowing links will fix it.” No — thin content is not a link problem; disavow does nothing for it.

Bottom line

Thin content is a value problem wearing a length costume. Judge each page on what it adds that a reader can’t get elsewhere — and if you’re publishing at scale, remember the damage is site-wide, not page-by-page. Improve, consolidate, or remove. That’s the whole job.

Add an expert note

Pin an expert quote

New person? Create their unclaimed profile at /admin/experts/ → Pin a quote first.