# Duplicate content

Duplicate content is the same or nearly the same content available at more than one URL. It wastes crawl, splits signals and can make Google show the wrong page. This page shows what to check, how to read it and how to fix it on a small site.

Updated 2026-09-14 · Source: https://porteur.ai/glossary/duplicate-content

## What duplicate content means for a small site

Duplicate content is identical or near-duplicate text or pages reachable at more than one URL, on your site or across sites. Think http and https, www and non-www, trailing slashes, parameters like ?ref or ?sort, print views, syndicated articles, or product descriptions copied from a manufacturer.

On a small site, duplicates dilute internal links and page authority. The wrong URL may rank or none will. Your /pricing could appear as /pricing, /pricing/, and /pricing?utm=twitter. You want one URL to win every time.

## Where duplicates come from

- Protocol and host: http vs https, www vs non-www.
- URL variants: trailing slash, uppercase vs lowercase, index.html on the end.
- Parameters: ?utm=, ?page=2, ?color=blue, session IDs.
- Content formats: print pages, AMP remnants, PDF versions of the same guide.
- Templates: category pagination repeating the same copy on /blog and /blog/page/2.
- Reused text: manufacturer product descriptions, syndicated posts, partner pages copying your copy.

## What Google does about duplicate content

Google says there is no duplicate content penalty. It chooses one URL as canonical and folds the rest. That choice can be wrong for your business page.

Deliberate duplication to manipulate ranking is different, that sits under spam policies. For normal site issues, your job is to help Google pick the right URL every time.

## How to find and measure it

- Search Console: check Pages, look for Duplicate, without user-selected canonical. Open a few examples and note the chosen canonical.
- URL Inspection: paste a duplicate URL, confirm the Google-selected canonical and the user-declared canonical.
- Your site: click key pages and copy their links. Check if your nav mixes /pricing and /pricing/. Fix internal links first.
- Server: request http and https, and www and non-www for your home page. Only one should load, the others should redirect 301.
- Parameters: list known parameters, like utm_source. Decide which change content. The rest are duplicates.
- Content sampling: pick five product pages. Search for a sentence in quotes. If many sites have it, expect Google to canonicalise to another source.

A fixed site shows one clean URL in Search Console as the canonical for each page, and your internal links all match it.

## How to fix and prevent duplicates

1. **Set one canonical home** Pick https and either www or non-www. Redirect the others with 301. Update settings at your host and CDN. Test both versions.
2. **Use rel=canonical on variants** On pages that must exist, add a canonical tag to the preferred URL. Example: /guides/print points to /guides as canonical.
3. **Redirect when you consolidate** If you remove /pricing/ in favour of /pricing, 301 redirect the old to the new. Update all internal links.
4. **Keep internal links consistent** Link to one version only. Do not link with and without trailing slash on the same site.
5. **Noindex versions for people** For pages that must exist but should not rank, like ?print or filtered lists, add a noindex. Keep them crawlable so canonical and noindex can be seen.
6. **Reduce reused text** Rewrite manufacturer blurbs. Change titles, intros and key specs. Make your product pages distinct.

Recheck the worst examples in URL Inspection after changes. You are done when Google-selected and user-declared canonicals match.

> Do not change URLs without a redirect map. Every old URL must 301 to the new preferred URL.

## Questions

### What does duplicate content mean?

It means the same or very similar content is accessible at more than one URL. It can be within one site, like /pricing and /pricing/, or across sites, like a syndicated article.

### Does duplicate content hurt SEO?

There is no specific duplicate content penalty. The risk is Google picks the wrong URL as canonical or merges signals in a way that costs you clicks. Fixing it helps the right page show.

### How do I find duplicate content on my site?

Start with Search Console, look for Duplicate, without user-selected canonical. Inspect a few URLs, then check http vs https, www vs non-www, trailing slashes and parameters. Sample product pages for copied text.

### How do you fix duplicate content?

Pick one URL as the canonical, redirect the others, keep internal links consistent, and add noindex to variants that must exist. Rewrite reused text so key pages are unique.

### Is a duplicate content checker necessary?

You can find most duplicates with Search Console and a manual crawl of your own links. A checker helps at scale, but start by fixing host, protocol, slashes and parameters.

### Should I use noindex or canonical on parameter pages?

If the parameter does not change the content, prefer canonical to the clean URL. If the page must exist but should not rank, use noindex. Test a few in URL Inspection and keep it consistent.

## Read next

- [Canonical tags: what they do and the mistakes that cost rankings](https://porteur.ai/guides/canonical-tag): What a canonical tag does, when to use one, the mistakes that cost rankings, and how to check and fix canonicals on a small site.
- [Canonical URL](https://porteur.ai/glossary/canonical-url): A canonical URL names the original version of a page. Use it to handle parameters and duplicates, avoid split signals and keep the right page indexed.
- [Duplicate without user-selected canonical: the fix in three steps](https://porteur.ai/guides/duplicate-without-user-selected-canonical): What the Search Console status means, where duplicates come from, how to pick the canonical, and the three steps to clear it for small product sites.
- [How to use the URL Inspection tool in Search Console](https://porteur.ai/guides/url-inspection-tool): Read each panel, run Test live URL, and know when to request indexing. Fix new pages, dropped pages, and canonicals Google ignores.
- [www or non-www: pick one, redirect the other](https://porteur.ai/guides/www-vs-non-www): Pick www or non-www, there is no SEO gain either way. 301 redirect every path to the chosen host, set canonicals, and verify a Domain property.
- [Trailing slash: two URLs for one page unless you decide](https://porteur.ai/guides/trailing-slash-seo): Google sees /page and /page/ as different. Pick one form, redirect the other, and keep links and your sitemap consistent. Here is how to do it.

Paste a URL and get a free check that spots duplicate versions and shows how rivals handle the same pages, in about thirty seconds. Free check: https://porteur.ai/
