Duplicate content
Duplicate content is the same or nearly the same content available at more than one URL. It wastes crawl, splits signals and can make Google show the wrong page. This page shows what to check, how to read it and how to fix it on a small site.
By Théophile Louvart, founder of Porteur · Updated 14 September 2026 · Markdown
What duplicate content means for a small site
Duplicate content is identical or near-duplicate text or pages reachable at more than one URL, on your site or across sites. Think http and https, www and non-www, trailing slashes, parameters like ?ref or ?sort, print views, syndicated articles, or product descriptions copied from a manufacturer.
On a small site, duplicates dilute internal links and page authority. The wrong URL may rank or none will. Your /pricing could appear as /pricing, /pricing/, and /pricing?utm=twitter. You want one URL to win every time.
Where duplicates come from
- Protocol and host: http vs https, www vs non-www.
- URL variants: trailing slash, uppercase vs lowercase, index.html on the end.
- Parameters: ?utm=, ?page=2, ?color=blue, session IDs.
- Content formats: print pages, AMP remnants, PDF versions of the same guide.
- Templates: category pagination repeating the same copy on /blog and /blog/page/2.
- Reused text: manufacturer product descriptions, syndicated posts, partner pages copying your copy.
What Google does about duplicate content
Google says there is no duplicate content penalty. It chooses one URL as canonical and folds the rest. That choice can be wrong for your business page.
Deliberate duplication to manipulate ranking is different, that sits under spam policies. For normal site issues, your job is to help Google pick the right URL every time.
How to find and measure it
- Search Console: check Pages, look for Duplicate, without user-selected canonical. Open a few examples and note the chosen canonical.
- URL Inspection: paste a duplicate URL, confirm the Google-selected canonical and the user-declared canonical.
- Your site: click key pages and copy their links. Check if your nav mixes /pricing and /pricing/. Fix internal links first.
- Server: request http and https, and www and non-www for your home page. Only one should load, the others should redirect 301.
- Parameters: list known parameters, like utm_source. Decide which change content. The rest are duplicates.
- Content sampling: pick five product pages. Search for a sentence in quotes. If many sites have it, expect Google to canonicalise to another source.
A fixed site shows one clean URL in Search Console as the canonical for each page, and your internal links all match it.
How to fix and prevent duplicates
Set one canonical home
Pick https and either www or non-www. Redirect the others with 301. Update settings at your host and CDN. Test both versions.
Use rel=canonical on variants
On pages that must exist, add a canonical tag to the preferred URL. Example: /guides/print points to /guides as canonical.
Redirect when you consolidate
If you remove /pricing/ in favour of /pricing, 301 redirect the old to the new. Update all internal links.
Keep internal links consistent
Link to one version only. Do not link with and without trailing slash on the same site.
Noindex versions for people
For pages that must exist but should not rank, like ?print or filtered lists, add a noindex. Keep them crawlable so canonical and noindex can be seen.
Reduce reused text
Rewrite manufacturer blurbs. Change titles, intros and key specs. Make your product pages distinct.
Recheck the worst examples in URL Inspection after changes. You are done when Google-selected and user-declared canonicals match.
Questions
It means the same or very similar content is accessible at more than one URL. It can be within one site, like /pricing and /pricing/, or across sites, like a syndicated article.
There is no specific duplicate content penalty. The risk is Google picks the wrong URL as canonical or merges signals in a way that costs you clicks. Fixing it helps the right page show.
Start with Search Console, look for Duplicate, without user-selected canonical. Inspect a few URLs, then check http vs https, www vs non-www, trailing slashes and parameters. Sample product pages for copied text.
Pick one URL as the canonical, redirect the others, keep internal links consistent, and add noindex to variants that must exist. Rewrite reused text so key pages are unique.
You can find most duplicates with Search Console and a manual crawl of your own links. A checker helps at scale, but start by fixing host, protocol, slashes and parameters.
If the parameter does not change the content, prefer canonical to the clean URL. If the page must exist but should not rank, use noindex. Test a few in URL Inspection and keep it consistent.
Sources
Check my site, free
Paste a URL and get a free check that spots duplicate versions and shows how rivals handle the same pages, in about thirty seconds.
- Free check, no card
- Read-only, your own accounts
- Readable by your agent
Read next
- GuideCanonical tags: what they do and the mistakes that cost rankings
- GlossaryCanonical URL
- GuideDuplicate without user-selected canonical: the fix in three steps
- GuideHow to use the URL Inspection tool in Search Console
- Guidewww or non-www: pick one, redirect the other
- GuideTrailing slash: two URLs for one page unless you decide
- GuideHow to fix duplicate content, case by case
- GuideContent decay: pages that earned more before, and the refresh