Canonicalization - Some advice needed :)
-
Hi guys,
To be honest, it's a little bit embarrassing to throw out this question but it's one of the weakest points of knowledge at the moment for me.
I've tried to get a grasp of canonical URLs and what it all means. From my understanding, it's informing Google which page to take into consideration when there's the possibility for duplicate content. Right?
However, with the site I'm working on I'm not sure if it would be worth putting site-wide and the impact it would have.
Site I'm working on - http://bit.ly/N7eew7
With the nature of the site, there would be a lot of duplicated content as there's the possibility that several properties listed could have a similar address due to being in the same building etc.
From what I can see, no canonical URL was setup on the homepage.
The other variations of the homepage URL are 301 redirecting to thee http:/www. version.
Can someone explain it all to me in simple terms? Honestly believe that I'm getting more confused by the minute.
Thanks guys for your patience
-
Seems like Matt and Marcus have you on the right track. With a real-estate site, duplicates and near-duplicates are very common, since you're adding and removing properties all the time and there are many search options and categories. I do agree that search-friendly URLs, long-term, where each property has a fixed URL, are definitely the best bet. In the meantime, though, a solid canonical structure helps a lot.
Ease into it - don't go sitewide in one fell swoop without a plan, unless you're having clear ranking problems. Start with your biggest problem areas, monitor/measure, and work from there. You can always check for indexed duplicates by running a Google search like:
site:daft.ie intitle:"176 Rathgar Road"
In this case, I'm not seeing any index issues, although I think Matt's concerns are valid.
I'd also consider rel=prev/next for search results pages, as that can help focus Google, too. Again, take it one step at a time and start with the biggest problems. It'll mitigate your risk all around.
-
What's everyones opinion canonicali URL being setup site-wide?
-
Hey, as per the email, it is exactly as above.
We can check the two versions of the URLs.
Confirm they both have the same canonical URL
then check both URLs using the info:URL command in Google to verify that in both instances, with and without final slash, the URL returned as indexed includes the final slash as per the canonical.
Any problems, give me a shout!
Marcus -
Hi Marcus thanks for your help so far. I've emailed you my URL's for a better look at the issue I'm facing.
-
Hi Antonio,
I hope you're well and not pulling your hair out in frustration just yet.
There are a few factors that you need to consider before making a decision on this:
1. Would changing the URL of the post give more traffic through the search engine than you are currently getting?
2. How would this impact the existing links that have been built to the original URL.
Remember that if you are going to change the URL of a page, this will just look like a new webpage to Google. All of the Facebook likes, Google+ +1's, links, etc will be going to the previous URL. Not only that, if you do a 301 redirect to the new URL, you will only transfer some of the link juice that you have made.
URL changes really should be a last resort and need to be thought out properly at the start of the webpage creation. In the case of Mark (above), I have recommended that he change the URLs because they are all dynamic and the benefit of changing these pages vs not, wins.
Let me know the URL of the page in question and I will take a look and tell you what I think.
Matt.
-
Hello Mathew and Mark congrats for the great support and highlights.
In the light of what you are explaning here could you please supoport me in this question concerning Canonical or 301 redirect? My issue is in terms of SEO when doing canolical.
I have a page with a long post title and url path name (more than 70 caracters and 115). This page has many visits but I am changing the SEO website structure according to SEOMOz and forums guidelines for the length names so: I WILL CREATE A DUPLICATE PAGE WITH THE SAME INFO.
This issue has been marked as an issue in the SEO tools, for long names>70 and url path names>115
My question is which option should I use and you would recommend me?
1. OPTION 1: Ideally I would like to keep the old post, so I should use the canonical tag, but my main concern is if the search engines in terms of SEO, even the canonical has been done, will penalise my SEO as there is still a post with bad SEO optimising, or if this is not the case because I already used the canonical. The duplicate content would still exist!
2. OPTION 2: Eliminate the post and redirection 301 to the new page to keep the juice.
I would prefer option 1, as I keep both post and page, but only if searchengines do not penalise my SEO as they detect a long post name and url path name.
Thank you very much for the help,
Antonio
-
Hi Matthew,
Thanks very much for your explanation. I think I get to understand it better now
Many thanks,
Christian
-
Will do - cheers Matthew
I'll probably take you up on that offer.
-
No problem.
I think the URLs should be the primary focus, and if you need any help on this, feel free to drop me a private message, etc and I will help you out.
Matt.
-
Hi Matthew, thanks for chipping in.
At the moment we do have canonical URLs setup for property listings such as the example you given above.
We'll still be going ahead with cleaning up the URL structure and ensuring categories following the correct practice as well.
-
Hi Christian,
No, this wouldn't be the case because what you are telling Google there is that "http://www.example.co.uk/properties/search" is the EXACT SAME page as the "/properties/search?page=1&commercialListingType=lease&propertyType=commercial/properties/search?page=1&commercialListingType=buy&propertyType=retail/" page.
For the likes of just search pages, you don't need to have canonical URLs because they are just dynamically generated search pages. Where you DO NEED canonical URLs is on the likes of category pages, product pages, etc.
So, in the case of Mark's website, the individual property listing pages (e.g, http://www.daft.ie/searchshortterm.daft?id=23606) need to have a canonical link because you could get to this page that has the EXACT SAME content with a similar URL (i don't know another URL to give the example here but a made up example could be http://www.daft.ie/searchshortterm.daft?id=23606keyword=dublin).
This is why you should have search engine friendly URLs to make it easy to understand which page is which. So having http://www.daft.ie/short-term/dublin/176-rathgar-road-apartment/ as the URL instead of http://www.daft.ie/searchshortterm.daft?id=23606 can make life a lot easier.
Has this helped to clear things up a bit?
Matt.
-
Hard to tell for 100% without the proper URLs but I don't think so.
You have one page that works on two different URLs. The page has a canonical tag showing that the http://www.mysite.com/product-a/ is the correct version.
So, in Googles eyes:
http://www.mysite.com/product-a/
http://www.mysite.com/product-aAre both pointing to:
http://www.mysite.com/product-a/
Due to the tag:
<link < span="">href="http://www.mysite.com/product-a/" rel="canonical" /> </link <>
There could be a bit more to this picture, if you don't want to post a link on here drop me an email to marcus@bowlerhat.co.uk and ill double check for you.
In an ideal world I would want consistency between URL's, site links and trailing slashes. I.E. If the page resolves on:
http://www.mysite.com/product-a
But is canonicalised to
http://www.mysite.com/product-a/
I would want a 301 from
http://www.mysite.com/product-a
to
http://www.mysite.com/product-a/
and all internal links to point to
http://www.mysite.com/product-a/
That's probably made it more confusing but in essence, nope, I think you are fine.
Cheers
Marcus
-
Hi Marcus
So here's what I've done...
So I've navigated like so:
Campaign>Crawl Diagnostics>Errors (68)>Duplicate Page Content Errors (61)Once this page loads all of the links, I've clicked on one of the links and it shows
1 Error
X Duplicate Page Content
Read MoreClicked on Read More then on the number 2 link that shows under the heading of Other URLs
This displays my two urls:
http://www.mysite.com/product-a/
http://www.mysite.com/product-aWhen I navigate to this page and view the source code I can see the following code:
href="http://www.mysite.com/product-a/" rel="canonical" />So I'm confused, do I have a duplicate content problem or not?
NB If I remove the trailing slash from my url it will show the same page. It does not do a redirect to the url with the slash. (I've highlighted this to Hubspot and they have said that it is not a problem?)
-
I don't believe that SEOMoz reports cover canonicalised links.
Simple test:
- Grab one page that has duplicate problems according to the report
- grab all duplicates from the spreadsheet
- Check the canonical on all
Mark - this is the same problem you will run into that I was trying to highlight above.
Marcus
-
I'm trialling seoMoz at the moment and so far I have 61 duplicate content crawl errors showing in one of my campaigns. This has sent me running to my CMS provider (Hubspot) to query this.
They've advised me that they automatically sort out canonicalisation.So I'm left in a state of not knowing where to focus.
Are Hubspot wrong or are the seoMoz reports broken?
-
Hi Christian,
That's a really good question - Can anyone shed any light on this one?
Personally I would have made the URL you mentioned be the canonical one.
But seeing I'm here asking for advice on it, maybe someone else would be better placed to help.
-
Well, you know, my dear old mother used to say an ounce of SEO prevention is worth a pound of SEO cure. Catch you later Mark.
-
Hi Mark and Marcus,
Sorry for jumping in your discussion; if i have URLs like below:
/properties/search?page=1&commercialListingType=lease&propertyType=commercial
/properties/search?page=1&commercialListingType=buy&propertyType=retail
does this mean that my canonical will be:
?
Many thanks for your help.
~Christian
-
Thanks Marcus - Agreed
Once URL structure has been improved, I will look into ensuring that specific property pages have canonical URLs and all relevant categories are appropriate setup as well.
Quite a bit of work to do but it should be worth it in the long term for the business.
-
Hi Mark,
No problem.
Yes, you are correct to assume that. For each of the property listings you would need to do this (just like the example that Marcus has given below).
I think that all areas of the website should really conform to these search engine friendly URLs. It may take quite a bit of time, but it will help you avoid a lot of issues in the future (which I can guarantee you would have).
Matt.
-
Yep, for sure, just beware it may still report duplication problems after you add the canonical URL so you will need to give it a manual once over. This is 100% worth doing though.
Marcus
-
Hi Marcus,
Just problems with the Moz tools.
We haven't been affected at all by any algorithm changes so far.
I still think it would be best to follow best practice going forward. I've just began work on this site and want to get to the root of any underlying problems.
Cheers,
Mark
-
Hey Mark
Are you having real world issues or just problems within the Moz tools?
I have feeling they don't factor canonicalisation at the moment (which sucks a bit) so you will do well to export the report to a spreadsheet and check them off manually.
Glad it was helpful!
Marcus
-
Marcus, thank you for giving such clear examples to me. It's a great help.
I'm a little bit embarrassed by the fact that it was causing such confusion up until now but it's clear to me now what needs to be changed.
With SEOMoz Campaign setup for the site, we have been receiving many duplicate content errors.
Hopefully the use of correct canonical URLs should help to eliminate many of the problems we have been having.
-
Hi Matt,
Thanks for the advice
Optimization of the URL structure is certainly something which I'm focusing on at the moment.
Taking on-board what you have mentioned, with the URL structure replaced, I presume that similar canonicals would need to be setup on each property listing to avoid duplicate content?
Do you think it's an issue which I should look into for other areas of the site as well?
Apologies for my questions. As you can guess, I'm trying to get to the root of any issues we're having with duplicate content.
Many thanks,
Mark
-
Hey Mark
In simple terms, the canonical URL exists as a suggestion to Google that a page may have various URLs or that various URLs may contain similar or near duplicate content.
For instance:
Lets say we have a list of properties in Birmingham, UK and that we have 3 pages showing that list of properties - the first by date order, the second by price high to low, the third by price low to high.
- http://www.example.co.uk/birmingham/properties.php
- http://www.example.co.uk/birmingham/properties.php?sort=hightolow
- http://www.example.co.uk/birmingham/properties.php?sort=lowtohigh
This is a perfect time to use the canonical URL as the content is the same, it is just jiggled around a little so all of these would set the default page as the canonical.
default page: http://www.example.co.uk/birmingham/properties.php
So, all pages would have this tag:
Then, Google knows that from a search and indexation perspective, they can return the one main version of this page and the others are just the same thing jumbled around a bit.
This is also a good, solid overview with a video and a basic explanation:
http://support.google.com/webmasters/bin/answer.py?hl=en&answer=139394
Hope that helps!
Marcus -
Hi Mark,
I hope you're well.
Basically, the canonical tag is used to let Google know which URL it should refer to as the original source of the page content. So, if you had the following URLs that all go to the homepage:
www.domain.com/
www.domain.com/index.php
www.domain.com/home/Then Google could crawl each of these pages and identify them as three different pages all with the same content. This could say to them that there is duplicate content on the site (which is not good). Usually with the homepage Google is intelligent enough to understand that there is just one page and the /index.php for example isn't a duplicate.
The problem that you do face, especially on the site that you are optimising, is with the different pages that have information on the lettings, etc (i.e. your product pages). For example, if you look at the following URL on your website:
http://www.daft.ie/searchshortterm.daft?id=23606
This is when you go through to the short-term searches and then I find the '176 Rathgar Road' apartment. Due to the dynamically generated URL (search.shortterm.daft?id=23606) I can gather that there would be several ways to get to this page with a different URL. My first suggestion would be to set up Search Engine Friendly URLs, for example, instead of having 'http://www.daft.ie/searchshortterm.daft?id=23606', it would be:
http://www.daft.ie/short-term/dublin/176-rathgar-road-apartment/
This way you could clearly optimise the page on Google search and have the canonical link to the page as:
href="http://www.daft.ie/short-term/dublin/176-rathgar-road-apartment.html" rel="canonical" />
This would improve the SEO performance on the website and avoid duplicate content issues.
I hope this helps, but if you need any more info then just let me know.
Matt.
Got a burning SEO question?
Subscribe to Moz Pro to gain full access to Q&A, answer questions, and ask your own.
Browse Questions
Explore more categories
-
Moz Tools
Chat with the community about the Moz tools.
-
SEO Tactics
Discuss the SEO process with fellow marketers
-
Community
Discuss industry events, jobs, and news!
-
Digital Marketing
Chat about tactics outside of SEO
-
Research & Trends
Dive into research and trends in the search industry.
-
Support
Connect on product support and feature requests.
Related Questions
-
Moving Blog and 301 Redirect Advice
Hello Moz Community, We recently moved our blog from its own domain to a directory on our website. We do not plan on moving over all the old blog posts because a majority most of them are based on events or time-sensitive information that has passed. We need advice on what to do with all of the old blog URL's? Should we just 301 all of them to the new blog directory on our website (www.domain.com/blog)? Should we take the time to move over all the old blog content and put the appropriate 301's in place? Any and all advice would be greatly appreciated. Thank you in advance.
Technical SEO | | All-Star-Vacation-Homes
Best,
Rich0 -
Starting a Blog and URL Structure Advice
Hello SEOmoz Community, We are going to start a blog on our website and have a slight dilemma. Our site is a .Net site and the blog platform we've chosen (BlogEngine) only allows us to use the following url structure: www.domain.com/blog/post/post-name. We've looked at other .Net blog software and this one meets all of our needs except for the ideal URL Structure. We would like to remove the /post/ directory; however have not technically found a way to do it. We wanted to get some opinions on whether or not we should just start with this URL structure and not worry about the extra directory, or work to find another solution that eliminates this extra directory. Ideally we want to keep the posts as close to the root as possible for link juice distribution, and the extra directory could get in the way. Also, if anyone has any advice on a more flexible .Net blog platform, suggestions would be greatly appreciated. We thank you so much in advance for your time and help.
Technical SEO | | All-Star-Vacation-Homes0 -
Is rel=canonical needed for URLs with Google Analytics query strings?
If a page URL has Google Analytics query strings, does the page need a canonical tag? e.g., something.com/?utm_source=newsletter&utm_campaign=mar-2013-nsl I have rel=canonical on all our pages because some of them will be accessed via URLs that have non-Google strings. The strings are only for marketing purposes, not for identifying a specific page to display. e.g., something.com/?source=acme Should I only implement the canonical tag on the pages that might have non-Google marketing strings in the URL?
Technical SEO | | WayneBlankenbeckler0 -
Wordpress & use of 'www' vs not for webmaster tools - explanation needed
I am having a hard time understanding the issue of canonization of site pages, specifically in regards to the 'www' or 'non-www' versions of a site. And specifically in regards to wordpress. I can see that it doesn't matter whether you type in 'www' or not in the url for a wordpress site, what is going on in the back end that allows this? When I link up to google webmaster tools, should i use www or not? thanks for any help d
Technical SEO | | dnaynay0 -
Do I need a 301 redirect on htaccess if Apache is already configured to serve?
Apache is set up to serve both www and non-www versions the same content. Do I still need to put a 301 redirect in the htaccess file?
Technical SEO | | Ocularis0 -
Wordpress site, combine Blog without hurting SEO - Need Expert Advice
Hi, I come from the old html days of Frontpage and then moved to Dreamweaver. I first worked with Wordpress at version 2.7 and was not all that impressed, but then recently I worked in the new version and was extremely impressed. So my knowledge of Wordpress is VERY limited and plan to build future sites with it. I need to know the best way to solve an issue for a customer. The client is http://www.nextgenrestoration.com/ Site was built years ago with Frontpage. The popularity of Blogs was hot so someone told them that if they add new content it would be better to use a blog, so they added a blog. So you have the following: www.nextgenrestoration.com (main site) then they installed wordpress in a folder (blog) www.nextgenrestoration.com/blog Original person that built the site quit. New person took over and said the main site needed to changed to Wordpress because they did not have Frontpage and all they knew was Wordpress. Main site was converted to Wordpress. They wanted to keep the original design so they did not use a stock template, they just built it with their design. I guess from looking at the Editor, they manually went in and put the design in to match. Now.. this last month, the person that had changed
Technical SEO | | Force7
the site to Wordpress quit. So I got involved because the new person they hired could not add content to the main website. If you add a page, it does not show up, you have to manually go in the php and add the link to the category. The new person knows how to use Wordpress but she knows nothing about PHP so is lost when it comes to manually adding content to the site. Here was my Thoughts. The main site needs to be rebuilt in a stock template so it automatically creates new pages, blog posts. I have to make sure that if we change the
main website that we could keep all the same links and page names. The girl
that built the site, if you hover over the links that she put it under ‘florida’,
that must be a category. But we would need to keep the same page names. I know
we could do a 301 redirect but this guy cannot lose traffic. He is already down
in hits after the last Panda update. My thought was, rebuild the main site in a stock template so
someone can actually add content easily to the site. Also build a new blog
section so it all matches. (personally the existing design looks old and dated and needs updating) If you look at the site now. The blog looks totally
different and it is not helping if a customer comes to the blog but cannot see
the navigation for the whole site. My thought was to just leave the old blog, it has a LOT of backlinks. But just add a new blog to the main site and all new content goes there. The old blog would stay just make sure we did build in some call to action so it sends them to the main site. Also, we found we cannot create a Blog on the
wordpress we have installed in the main directory. I am guessing because it
wants to name it /blog? I want to be sure we give this client the best advice on what to do without
hurting his existing seo and traffic. As you can tell, I am not qualified to really give the best advice since I am so new to Wordpress. This is a small company that really needs some help. Thanks in advance for your time! Force70 -
Advice on strange URL problem
I'm considering doing some pro bono work for a local non-profit and upon initial review they have a number of serious issues but there is one in particular I'd like to check my thinking on. The developer who set up the site some years ago implemented a javascript redirect on their root domain so that it redirects to: http://domain.com/wordpress This is wrong for all kinds of reasons and I want to recommend eliminating this redirect and getting rid of the 'wordpress' part of the path altogether. However, the site is quite established with good PR and they would take a hit by changing the path. I'd do 301 redirects to the new URLs that would not have 'wordpress' in the path in addition to other remediation. My question - is my thinking here good? It's worth it, right? The other option is just get rid of the weird redirect and keep 'wordpress' in the path but this seems unacceptable to me. Any opinions?
Technical SEO | | friendlymachine0