Your homepage might live at four addresses right now: with and without www, with HTTP and HTTPS. To a search engine those look like four competing pages splitting your ranking power — unless a canonical tag tells it they're one. Here's how that works.
Where duplicates come from
Duplicate URLs are rarely created deliberately. They accumulate: protocol variants (http/https), host variants (www/apex), trailing slashes, UTM tracking parameters, session IDs, pagination, print views, and faceted filters that generate thousands of near-identical combinations. Each variant can get crawled, indexed, and linked separately — diluting the signals that determine rankings.
Related reading: How to Remain Valuable When Intelligence Becomes Cheap — a 224-page practical book on staying valuable as intelligence gets cheap. $3.84. Read it on Gumroad →
What the canonical tag does
A canonical tag in the page's <head> names the one true URL for that content:
<link rel="canonical" href="https://example.com/blog/my-post" />
It's a consolidation signal, not a redirect: crawlers still fetch the duplicate, but they credit the canonical version with the ranking signals — links, relevance, authority. Every page should carry a self-referencing canonical (pointing to itself), so that any parameterized or miscased variant of its URL automatically defers to the clean one.
Mistakes that break canonicals
- Chains and loops. If A canonicalizes to B and B canonicalizes to C, search engines may ignore both hops. Point every duplicate directly at the final URL.
- Canonical to a redirect. The canonical target should return 200, not bounce through redirects. Resolve the chain first.
- Conflicting signals. A page canonicalized to URL X but included in the sitemap as URL Y sends mixed messages. Keep sitemap, canonicals, and internal links in agreement.
- Cross-domain overreach. Pointing your canonical at someone else's page (common with syndicated content) hands them your ranking credit. Usually that's exactly backwards.
Audit yours
Canonical issues are silent — nothing looks broken to visitors while rankings quietly leak. The canonical URL checker inspects a page's canonical setup, and pairing it with a crawl of your sitemap URLs catches the conflicts above. For the crawler's-eye view of how these signals get discovered, see how search engines crawl and index websites.