Duplicate content
Duplicate content is the same or very similar content available at multiple URLs, which splits ranking signals and forces search engines to pick one.
Duplicate content is substantially the same content reachable at more than one URL. It is rarely a penalty, but it is a dilution problem: search engines must choose which version to rank, and the links and relevance that should concentrate on one URL get split across several, weakening all of them.
How it works
Duplicates arise structurally far more than by copying: URL parameters, faceted combinations, HTTP and HTTPS, www and non-www, trailing-slash variants, print and AMP versions, and syndication. When Google finds duplicates, it clusters them and picks a canonical to represent the group; the others get filtered. You control that choice with canonical tags, 301 redirects, consistent internal linking, and parameter handling. The goal is one authoritative URL per piece of content, with every signal pointing to it.
Where it applies
- Ecommerce — the same product across categories, plus parameter and facet variants.
- Publishing and syndication — articles republished on partners or aggregators that need a canonical home.
- Migrations — old and new URLs coexisting and competing during a transition.
What matters from each seat
- In-house — enforce self-referencing canonicals and consistent internal links so signals consolidate.
- Engineering — normalize protocol, host, case, and trailing slashes at the platform level.
- Agency and consulting — a duplicate-content audit is a fast, high-impact early win, especially after a migration.
Have a search problem worth solving?
Bring the migration, the ranking drop, or the question you cannot get a straight answer on. One call tells us both whether I can help.