Skip to content
weballin

weballin DIGIWIKI

Duplicate Content

Duplicate content is identical or very similar material available at multiple URLs. Ordinary technical duplication does not automatically mean a penalty; clarify the preferred URL and the purpose of each version.

weballin

Print views, tracking parameters or different hostnames can expose the same material. First compare the body and determine whether each URL needs to remain available.

Consider canonical annotations when duplicate versions must remain accessible, or redirects when a retired URL has an equivalent destination. Align internal links and sitemaps with the preferred URL. Noindex excludes content from indexing and is not the recommended substitute for canonical consolidation.

Do not assume that a scraper’s copying automatically penalizes the original author. Use crawling and body comparisons to investigate duplicates. Use Search Console for indexing status and Google-selected canonicals, and a crawler or similar method to compare bodies across pages.

References

Frequently asked questions

Do matching titles prove pages are duplicates?

No. Compare the body, purpose and data. Distinct pages should nevertheless have titles that help readers tell them apart.