Thin content
Thin content is a page that offers little unique value to users, and at scale it can drag down how search engines judge a whole section or site.
Thin content is a page with little unique value — auto-generated combinations, near-duplicate variants, doorway pages, or content that adds nothing a searcher couldn’t get elsewhere. A single thin page is harmless; thousands of them signal low quality and can suppress the performance of otherwise strong pages around them.
Why it matters
Search engines assess quality partly at the site and section level, so a large volume of thin, indexable pages dilutes the overall quality signal and wastes crawl budget. Thin content typically appears where pages are generated mechanically: faceted-navigation combinations, tag archives, boilerplate location pages, and syndicated duplicates. The remedy is to consolidate, enrich, or de-index — keep the pages that serve real demand with genuine value, and remove the rest from the index with noindex or canonical tags. For answer engines, thin pages simply aren’t citable; there is nothing worth lifting.
Where it applies
- Ecommerce — facet and parameter URLs that spawn thin, near-duplicate category variants.
- Programmatic SEO — templated pages (locations, combinations) that must clear a real value bar to index.
- Publishing — tag and archive pages, or short posts that repeat what deeper pages already cover.
What matters from each seat
- In-house — set an indexation bar: a page earns indexing by adding unique value, or it stays out.
- Editorial — depth and originality, not word count, are what make a page not thin.
- Agency and consulting — a content-quality and index-bloat audit often finds thousands of pages worth pruning.
Have a search problem worth solving?
Bring the migration, the ranking drop, or the question you cannot get a straight answer on. One call tells us both whether I can help.