Your homepage might live at four addresses right now: with and without www, with HTTP and HTTPS. To a search engine those look like four competing pages splitting your ranking power — unless a canonical tag tells it they're one. Here's how that works.

Where duplicates come from

Duplicate URLs are rarely created deliberately. They accumulate: protocol variants (http/https), host variants (www/apex), trailing slashes, UTM tracking parameters, session IDs, pagination, print views, and faceted filters that generate thousands of near-identical combinations. Each variant can get crawled, indexed, and linked separately — diluting the signals that determine rankings.

Related reading: How to Remain Valuable When Intelligence Becomes Cheap — a 224-page practical book on staying valuable as intelligence gets cheap. $3.84. Read it on Gumroad →

What the canonical tag does

A canonical tag in the page's <head> names the one true URL for that content:

<link rel="canonical" href="https://example.com/blog/my-post" />

It's a consolidation signal, not a redirect: crawlers still fetch the duplicate, but they credit the canonical version with the ranking signals — links, relevance, authority. Every page should carry a self-referencing canonical (pointing to itself), so that any parameterized or miscased variant of its URL automatically defers to the clean one.

Mistakes that break canonicals

  • Chains and loops. If A canonicalizes to B and B canonicalizes to C, search engines may ignore both hops. Point every duplicate directly at the final URL.
  • Canonical to a redirect. The canonical target should return 200, not bounce through redirects. Resolve the chain first.
  • Conflicting signals. A page canonicalized to URL X but included in the sitemap as URL Y sends mixed messages. Keep sitemap, canonicals, and internal links in agreement.
  • Cross-domain overreach. Pointing your canonical at someone else's page (common with syndicated content) hands them your ranking credit. Usually that's exactly backwards.

Audit yours

Canonical issues are silent — nothing looks broken to visitors while rankings quietly leak. The canonical URL checker inspects a page's canonical setup, and pairing it with a crawl of your sitemap URLs catches the conflicts above. For the crawler's-eye view of how these signals get discovered, see how search engines crawl and index websites.