Glossary · Technical SEO

Duplicate content

Also called: duplicate pages, near-duplicate content

Definition

Duplicate content is substantially the same content appearing at more than one URL, on the same site or across different sites.

Duplicate content explained

Most duplicate content is accidental and technical: HTTP and HTTPS versions, www and non-www, trailing slashes, URL parameters, printer-friendly pages, faceted filters, and products listed in several categories. Other duplication comes from content itself, such as manufacturer descriptions reused by many retailers or location pages that differ only by the city name.

Duplicate content is not usually a penalty issue. The problem is dilution and choice: search engines pick one version to show, which may not be the one you want, and links and signals are split across the copies. Large-scale duplication can also waste crawl budget and make a site look low in value.

Fixes depend on the cause: redirect duplicate hostnames and protocols to one version, set canonical tags on parameter and variant URLs, keep filter combinations out of the index, and write unique content for pages that deserve to rank on their own.

Example

Your site loads at both http and https, with and without www, giving every page four addresses. Redirecting all versions to https with www, and setting matching canonical tags, turns four competing copies of each page into one.

Why it matters

Consolidating duplicates makes it clear which URL should rank and concentrates all its signals there.

Related service

Technical SEO

Crawlability, indexing, structured data and rendering fixes so search engines and AI crawlers can read your site.

Technical SEO services

Published by Vidern, founded and led by Malhar Shah. Updated .

Find out what's holding your site back

Get a free SEO audit with prioritized fixes, delivered within 48 hours.