Duplicate Content
The same content reachable under several addressesWhen identical or near-identical text is reachable under several addresses — there's no penalty for that, but there is a selection that Google makes itself.
- One piece of content ends up under several addresses
- Google filters instead of penalizing
- One address is chosen as authoritative
- Authority concentrates when signals agree
What does Duplicate Content mean?
Duplicate content almost always happens unintentionally: the same page with and without www, with a session ID from a campaign, as a print version, as a filtered view in a shop, or because a text appears in a blog post and in a topic overview at the same time. A single piece of content ends up under several addresses that look identical to people and, to a search engine, initially look like several separate pages.
A persistent misunderstanding has attached itself to the term: the idea of a 'duplicate content penalty' that actively downgrades a website for duplicate content. Google explicitly contradicts this in its own documentation — there is no such penalty in the usual sense. What happens instead, Google calls filtering or canonicalization: of several equivalent addresses, one is selected for the results, the others are set aside, not penalized.
The search engine makes this selection algorithmically based on signals — internal linking, redirects, entries in the address directory, and the canonical tag, which serves as a hint about which address should be authoritative. Google mostly follows this hint, but not necessarily. With conflicting signals, an address other than the intended one can be chosen, without any error message pointing this out.
The real damage lies not in a downgrade but in dilution: if half the website links to the www variant and the other half to the variant without www, the authority a page has accumulated is split across two addresses instead of concentrating in one. On top of that, the search engine wastes crawl capacity on variants that won't be shown in the end anyway.
One exception adds to the confusion: in large-scale copying of other people's content across many domains, for example in spam networks, Google can actually intervene — but that runs through the policies against spam, not through a separate duplicate content rule. The ordinary case on your own website, which is what this entry is about, is different from that and considerably more harmless.
In practice that means: fix one address variant (with or without www, with or without a trailing slash), reflect that decision consistently in the canonical tag, route genuinely moved content through a redirect rather than a hint, and avoid publishing the same text twice on the same website without good reason. Details on the technical implementation of the hint are in the entry on the canonical tag; here the underlying rule is what matters most.
Check the numbers
For this entry we deliberately give no figure — there is no measurable 'penalty value' we could cite, because the penalty itself doesn't exist.
The point we can substantiate instead is a statement, not a metric: Google states plainly in its own documentation that 'in most cases' there is no duplicate content penalty, only automatic filtering. Inventing a figure for it, say an alleged ranking loss in percent, would be exactly the kind of claim we want to warn against here. What can be measured is the individual case — which address Google actually chose is shown in Google Search Console.
Source: Google Search Central, documentation on duplicate content and canonicalization. No penalty for duplicate content is described there.
Common questions
Will my website be penalized if a text appears twice?
Not in the usual sense. Google itself speaks of filtering rather than a penalty: of several identical or similar addresses, one is selected and shown, the others set aside. Damage tends to arise indirectly, through diluted authority and wasted crawl capacity.
Is setting the canonical tag enough to solve the problem?
In most cases, yes, because it tells the search engine which address should count. But it's a hint, not an instruction — with conflicting internal links or missing redirects, Google can still choose a different address.
A term missing? Send it to us
We explain every term calmly and without tech-speak — and tell you honestly what makes sense for you and what doesn't.