What is a Canonical URL? Definition & Best Practices
Definition
A canonical URL is the preferred version of a page that you want search engines to index and rank. It is declared with a <link rel="canonical" href="..."> tag in the HTML <head>, telling crawlers which URL is the master copy when several pages have identical or very similar content.
Why canonical URLs matter
- Consolidate ranking signals — link equity and relevance signals point to one URL instead of being split across duplicates.
- Prevent duplicate content — pagination, filter parameters, tracking parameters and session IDs can create dozens of URLs for the same content.
- Control which URL appears in search results — search engines display the canonical version, not the variants.
- Save crawl budget — Google focuses on the canonical version instead of crawling every duplicate.
Common canonical mistakes
- Canonicalizing to a 404 or redirect — points search engines at a broken page. Scan your site to detect canonicals pointing to non-200 URLs.
- Conflicting signals — using a canonical that contradicts hreflang, sitemap or internal linking.
- Self-referencing on every page with the wrong protocol (http vs https) or wrong domain (www vs non-www).
Canonical URLs are one of the most powerful tools for managing duplicate content. Combined with consistent internal linking and proper redirects, they keep your indexable surface clean and your ranking signals concentrated.
Related terms
Crawl Budget
How search engines allocate resources to crawl your site — and why broken links waste it.
hreflang
The HTML attribute that tells search engines which language version to serve.
Internal Link
Links between pages on the same domain — the backbone of site structure.
Link Equity
The ranking value passed through links — also known as "link juice".
www vs non-www
Choosing the canonical hostname — and why consistency matters more than the choice itself.
Link Rot
The gradual decay of links over time as pages move or disappear.