Canonicalization - Some advice needed :)
-
Hi guys,
To be honest, it's a little bit embarrassing to throw out this question but it's one of the weakest points of knowledge at the moment for me.
I've tried to get a grasp of canonical URLs and what it all means. From my understanding, it's informing Google which page to take into consideration when there's the possibility for duplicate content. Right?
However, with the site I'm working on I'm not sure if it would be worth putting site-wide and the impact it would have.
Site I'm working on - http://bit.ly/N7eew7
With the nature of the site, there would be a lot of duplicated content as there's the possibility that several properties listed could have a similar address due to being in the same building etc.
From what I can see, no canonical URL was setup on the homepage.
The other variations of the homepage URL are 301 redirecting to thee http:/www. version.
Can someone explain it all to me in simple terms? Honestly believe that I'm getting more confused by the minute.
Thanks guys for your patience
-
Seems like Matt and Marcus have you on the right track. With a real-estate site, duplicates and near-duplicates are very common, since you're adding and removing properties all the time and there are many search options and categories. I do agree that search-friendly URLs, long-term, where each property has a fixed URL, are definitely the best bet. In the meantime, though, a solid canonical structure helps a lot.
Ease into it - don't go sitewide in one fell swoop without a plan, unless you're having clear ranking problems. Start with your biggest problem areas, monitor/measure, and work from there. You can always check for indexed duplicates by running a Google search like:
site:daft.ie intitle:"176 Rathgar Road"
In this case, I'm not seeing any index issues, although I think Matt's concerns are valid.
I'd also consider rel=prev/next for search results pages, as that can help focus Google, too. Again, take it one step at a time and start with the biggest problems. It'll mitigate your risk all around.
-
What's everyones opinion canonicali URL being setup site-wide?
-
Hey, as per the email, it is exactly as above.
We can check the two versions of the URLs.
Confirm they both have the same canonical URL
then check both URLs using the info:URL command in Google to verify that in both instances, with and without final slash, the URL returned as indexed includes the final slash as per the canonical.
Any problems, give me a shout!
Marcus -
Hi Marcus thanks for your help so far. I've emailed you my URL's for a better look at the issue I'm facing.
-
Hi Antonio,
I hope you're well and not pulling your hair out in frustration just yet.
There are a few factors that you need to consider before making a decision on this:
1. Would changing the URL of the post give more traffic through the search engine than you are currently getting?
2. How would this impact the existing links that have been built to the original URL.
Remember that if you are going to change the URL of a page, this will just look like a new webpage to Google. All of the Facebook likes, Google+ +1's, links, etc will be going to the previous URL. Not only that, if you do a 301 redirect to the new URL, you will only transfer some of the link juice that you have made.
URL changes really should be a last resort and need to be thought out properly at the start of the webpage creation. In the case of Mark (above), I have recommended that he change the URLs because they are all dynamic and the benefit of changing these pages vs not, wins.
Let me know the URL of the page in question and I will take a look and tell you what I think.
Matt.
-
Hello Mathew and Mark congrats for the great support and highlights.
In the light of what you are explaning here could you please supoport me in this question concerning Canonical or 301 redirect? My issue is in terms of SEO when doing canolical.
I have a page with a long post title and url path name (more than 70 caracters and 115). This page has many visits but I am changing the SEO website structure according to SEOMOz and forums guidelines for the length names so: I WILL CREATE A DUPLICATE PAGE WITH THE SAME INFO.
This issue has been marked as an issue in the SEO tools, for long names>70 and url path names>115
My question is which option should I use and you would recommend me?
1. OPTION 1: Ideally I would like to keep the old post, so I should use the canonical tag, but my main concern is if the search engines in terms of SEO, even the canonical has been done, will penalise my SEO as there is still a post with bad SEO optimising, or if this is not the case because I already used the canonical. The duplicate content would still exist!
2. OPTION 2: Eliminate the post and redirection 301 to the new page to keep the juice.
I would prefer option 1, as I keep both post and page, but only if searchengines do not penalise my SEO as they detect a long post name and url path name.
Thank you very much for the help,
Antonio
-
Hi Matthew,
Thanks very much for your explanation. I think I get to understand it better now
Many thanks,
Christian
-
Will do - cheers Matthew
I'll probably take you up on that offer.
-
No problem.
I think the URLs should be the primary focus, and if you need any help on this, feel free to drop me a private message, etc and I will help you out.
Matt.
-
Hi Matthew, thanks for chipping in.
At the moment we do have canonical URLs setup for property listings such as the example you given above.
We'll still be going ahead with cleaning up the URL structure and ensuring categories following the correct practice as well.
-
Hi Christian,
No, this wouldn't be the case because what you are telling Google there is that "http://www.example.co.uk/properties/search" is the EXACT SAME page as the "/properties/search?page=1&commercialListingType=lease&propertyType=commercial/properties/search?page=1&commercialListingType=buy&propertyType=retail/" page.
For the likes of just search pages, you don't need to have canonical URLs because they are just dynamically generated search pages. Where you DO NEED canonical URLs is on the likes of category pages, product pages, etc.
So, in the case of Mark's website, the individual property listing pages (e.g, http://www.daft.ie/searchshortterm.daft?id=23606) need to have a canonical link because you could get to this page that has the EXACT SAME content with a similar URL (i don't know another URL to give the example here but a made up example could be http://www.daft.ie/searchshortterm.daft?id=23606keyword=dublin).
This is why you should have search engine friendly URLs to make it easy to understand which page is which. So having http://www.daft.ie/short-term/dublin/176-rathgar-road-apartment/ as the URL instead of http://www.daft.ie/searchshortterm.daft?id=23606 can make life a lot easier.
Has this helped to clear things up a bit?
Matt.
-
Hard to tell for 100% without the proper URLs but I don't think so.
You have one page that works on two different URLs. The page has a canonical tag showing that the http://www.mysite.com/product-a/ is the correct version.
So, in Googles eyes:
http://www.mysite.com/product-a/
http://www.mysite.com/product-aAre both pointing to:
http://www.mysite.com/product-a/
Due to the tag:
<link < span="">href="http://www.mysite.com/product-a/" rel="canonical" /> </link <>
There could be a bit more to this picture, if you don't want to post a link on here drop me an email to marcus@bowlerhat.co.uk and ill double check for you.
In an ideal world I would want consistency between URL's, site links and trailing slashes. I.E. If the page resolves on:
http://www.mysite.com/product-a
But is canonicalised to
http://www.mysite.com/product-a/
I would want a 301 from
http://www.mysite.com/product-a
to
http://www.mysite.com/product-a/
and all internal links to point to
http://www.mysite.com/product-a/
That's probably made it more confusing but in essence, nope, I think you are fine.
Cheers
Marcus
-
Hi Marcus
So here's what I've done...
So I've navigated like so:
Campaign>Crawl Diagnostics>Errors (68)>Duplicate Page Content Errors (61)Once this page loads all of the links, I've clicked on one of the links and it shows
1 Error
X Duplicate Page Content
Read MoreClicked on Read More then on the number 2 link that shows under the heading of Other URLs
This displays my two urls:
http://www.mysite.com/product-a/
http://www.mysite.com/product-aWhen I navigate to this page and view the source code I can see the following code:
href="http://www.mysite.com/product-a/" rel="canonical" />So I'm confused, do I have a duplicate content problem or not?
NB If I remove the trailing slash from my url it will show the same page. It does not do a redirect to the url with the slash. (I've highlighted this to Hubspot and they have said that it is not a problem?)
-
I don't believe that SEOMoz reports cover canonicalised links.
Simple test:
- Grab one page that has duplicate problems according to the report
- grab all duplicates from the spreadsheet
- Check the canonical on all
Mark - this is the same problem you will run into that I was trying to highlight above.
Marcus
-
I'm trialling seoMoz at the moment and so far I have 61 duplicate content crawl errors showing in one of my campaigns. This has sent me running to my CMS provider (Hubspot) to query this.
They've advised me that they automatically sort out canonicalisation.So I'm left in a state of not knowing where to focus.
Are Hubspot wrong or are the seoMoz reports broken?
-
Hi Christian,
That's a really good question - Can anyone shed any light on this one?
Personally I would have made the URL you mentioned be the canonical one.
But seeing I'm here asking for advice on it, maybe someone else would be better placed to help.
-
Well, you know, my dear old mother used to say an ounce of SEO prevention is worth a pound of SEO cure. Catch you later Mark.
-
Hi Mark and Marcus,
Sorry for jumping in your discussion; if i have URLs like below:
/properties/search?page=1&commercialListingType=lease&propertyType=commercial
/properties/search?page=1&commercialListingType=buy&propertyType=retail
does this mean that my canonical will be:
?
Many thanks for your help.
~Christian
-
Thanks Marcus - Agreed
Once URL structure has been improved, I will look into ensuring that specific property pages have canonical URLs and all relevant categories are appropriate setup as well.
Quite a bit of work to do but it should be worth it in the long term for the business.
-
Hi Mark,
No problem.
Yes, you are correct to assume that. For each of the property listings you would need to do this (just like the example that Marcus has given below).
I think that all areas of the website should really conform to these search engine friendly URLs. It may take quite a bit of time, but it will help you avoid a lot of issues in the future (which I can guarantee you would have).
Matt.
-
Yep, for sure, just beware it may still report duplication problems after you add the canonical URL so you will need to give it a manual once over. This is 100% worth doing though.
Marcus
-
Hi Marcus,
Just problems with the Moz tools.
We haven't been affected at all by any algorithm changes so far.
I still think it would be best to follow best practice going forward. I've just began work on this site and want to get to the root of any underlying problems.
Cheers,
Mark
-
Hey Mark
Are you having real world issues or just problems within the Moz tools?
I have feeling they don't factor canonicalisation at the moment (which sucks a bit) so you will do well to export the report to a spreadsheet and check them off manually.
Glad it was helpful!
Marcus
-
Marcus, thank you for giving such clear examples to me. It's a great help.
I'm a little bit embarrassed by the fact that it was causing such confusion up until now but it's clear to me now what needs to be changed.
With SEOMoz Campaign setup for the site, we have been receiving many duplicate content errors.
Hopefully the use of correct canonical URLs should help to eliminate many of the problems we have been having.
-
Hi Matt,
Thanks for the advice
Optimization of the URL structure is certainly something which I'm focusing on at the moment.
Taking on-board what you have mentioned, with the URL structure replaced, I presume that similar canonicals would need to be setup on each property listing to avoid duplicate content?
Do you think it's an issue which I should look into for other areas of the site as well?
Apologies for my questions. As you can guess, I'm trying to get to the root of any issues we're having with duplicate content.
Many thanks,
Mark
-
Hey Mark
In simple terms, the canonical URL exists as a suggestion to Google that a page may have various URLs or that various URLs may contain similar or near duplicate content.
For instance:
Lets say we have a list of properties in Birmingham, UK and that we have 3 pages showing that list of properties - the first by date order, the second by price high to low, the third by price low to high.
- http://www.example.co.uk/birmingham/properties.php
- http://www.example.co.uk/birmingham/properties.php?sort=hightolow
- http://www.example.co.uk/birmingham/properties.php?sort=lowtohigh
This is a perfect time to use the canonical URL as the content is the same, it is just jiggled around a little so all of these would set the default page as the canonical.
default page: http://www.example.co.uk/birmingham/properties.php
So, all pages would have this tag:
Then, Google knows that from a search and indexation perspective, they can return the one main version of this page and the others are just the same thing jumbled around a bit.
This is also a good, solid overview with a video and a basic explanation:
http://support.google.com/webmasters/bin/answer.py?hl=en&answer=139394
Hope that helps!
Marcus -
Hi Mark,
I hope you're well.
Basically, the canonical tag is used to let Google know which URL it should refer to as the original source of the page content. So, if you had the following URLs that all go to the homepage:
www.domain.com/
www.domain.com/index.php
www.domain.com/home/Then Google could crawl each of these pages and identify them as three different pages all with the same content. This could say to them that there is duplicate content on the site (which is not good). Usually with the homepage Google is intelligent enough to understand that there is just one page and the /index.php for example isn't a duplicate.
The problem that you do face, especially on the site that you are optimising, is with the different pages that have information on the lettings, etc (i.e. your product pages). For example, if you look at the following URL on your website:
http://www.daft.ie/searchshortterm.daft?id=23606
This is when you go through to the short-term searches and then I find the '176 Rathgar Road' apartment. Due to the dynamically generated URL (search.shortterm.daft?id=23606) I can gather that there would be several ways to get to this page with a different URL. My first suggestion would be to set up Search Engine Friendly URLs, for example, instead of having 'http://www.daft.ie/searchshortterm.daft?id=23606', it would be:
http://www.daft.ie/short-term/dublin/176-rathgar-road-apartment/
This way you could clearly optimise the page on Google search and have the canonical link to the page as:
href="http://www.daft.ie/short-term/dublin/176-rathgar-road-apartment.html" rel="canonical" />
This would improve the SEO performance on the website and avoid duplicate content issues.
I hope this helps, but if you need any more info then just let me know.
Matt.
Got a burning SEO question?
Subscribe to Moz Pro to gain full access to Q&A, answer questions, and ask your own.
Browse Questions
Explore more categories
-
Moz Tools
Chat with the community about the Moz tools.
-
SEO Tactics
Discuss the SEO process with fellow marketers
-
Community
Discuss industry events, jobs, and news!
-
Digital Marketing
Chat about tactics outside of SEO
-
Research & Trends
Dive into research and trends in the search industry.
-
Support
Connect on product support and feature requests.
Related Questions
-
New SEO manager needs help! Currently only about 15% of our live sitemap (~4 million url e-commerce site) is actually indexed in Google. What are best practices sitemaps for big sites with a lot of changing content?
In Google Search console 4,218,017 URLs submitted 402,035 URLs indexed what is the best way to troubleshoot? What is best guidance for sitemap indexation of large sites with a lot of changing content? view?usp=sharing
Technical SEO | | Hamish_TM1 -
Moving from http to https - what do I need to do in Google Search Console?
Hi all, I have moved my site from http to https. I current have two profiles in Google Search Console: http://mysite.com
Technical SEO | | Bee159
http://www.mysite.com Do I need to set up the same but with https and if so, what do I then do with the http profiles? Do I delete them? Or just remove the sitemaps? Confused.0 -
What are the things that need update? before Appying SEO
Hello Everyone, I need Start work on for this site www.tajsigma.com, I want to know Did i need Add any things for Good ranking. what are the things that need update in This site For SEO??
Technical SEO | | falguniinnovative
i know Basic Things robots, title, discription , But i want know Any Additional Things that I need to apply?? can anyone help, please Thanx in advance0 -
Wrapping my head around an e-commerce anchor filter issue, need help
I am having a hard time understanding how Google will deal with this scenario, I would love to hear what you guys think or suggest. Ok a category page on the site in question looks like this. http://makeupaddict.me/6-skin-care All fine and well, But a paginated page or a filtered category pages look like these http://makeupaddict.me/6-skin-care#/page-2 and http://makeupaddict.me/6-skin-care#/price-391-1217 From my understanding Google does not index an anchor without a shebang (#!), but that doesn't mean that they do not still crawl them, correct? That is where the issue comes in, since anchors are not indexed and dropped from the urls, when Google crawls a filtered or paginated page, it is getting different results. From the best of my understanding, and someone can correct me if I am wrong but an anchor is not passed in web languages like a querystring is. So if I am using php and land on http://makeupaddict.me/6-skin-care or http://makeupaddict.me/6-skin-care#/price-391-1217 and use something like .$_SERVER['SELF'] to get the url both pages will return http://makeupaddict.me/6-skin-care since the anchor is handled client side. With that being the case, is it imagined that Google uses that standard or is it thought they have a custom function that grabs the whole url anchor in all? Also if they are crawling the page with the anchor, but seeing it anchor less how are they handling the changing content?
Technical SEO | | LesleyPaone0 -
Do I need to 301 EVERY page?
I have a client who is consolidating multiple EMD domains into a single domain for SEO reasons and for practical reason, like not having to produce content and perform SEO for 20 domains. My question is this: Do I need to 301 every single page from these old EMD domains? I bill this client hourly and while I could take the time to write 301s for literally thousands of pages I feel that this might not be the best use of his money, that I could strategically 301 the landing pages that get traffic and then route everything else to the new root domain...thoughts? I've researched this and have not been able to hear a really strong opinion yet.
Technical SEO | | BrianJGomez0 -
Duplicate index.php/webpage pages on website. Help needed!
Hi Guys, Having a really frustrating problem with our website. It is a Joomla 1.7 site and we have some duplicate page issues. What is happening is that we have a webpage, lets say domain.com/webpage1 and then we also have domain.com/index.php/webpage1. Google is seeing these as duplicate pages and is causing me some real SEO problems. I have tried setting up a 301 redirect but it wn't let me redirect /index.php/webpage1 to /webpage1. Anyone have any ideas or plugins that can be used to sort this out? Any help will be really appreciated! Matt.
Technical SEO | | MatthewBarby0 -
Homepage canonicalized with trailing slash
We were told by a consultant that in SEO it is best practice to canonicalize our homepage URL with the trailing slash. What do you think about doing this for the homepage? Is it important for other site links to have the trailing slash as well
Technical SEO | | fibers0 -
Need advanced SEO help!
Hi guys, This is my last attempt to work out what is up with this site before it goes to the big Flipper in the sky (and even then I doubt it will make much more than £1!) This site was a successful site, then one day Google decided it didnt like it, and I have not had much joy with it for nearly a year now. I must admit I tried to forget about it for a while, but it has always been a thorn in my side due to the fact it used to be a nice little earner. I have SEOmoz crawled it and I cant find any issues that would cause such a severe penalty, I removed many of the affiliate links, clocked the rest of the affiliate links and tried numurous other ideas, but now, as a last ditch attempt I am looking for some help! I tried to avoid the typical thin affiliate site by adding relevant content, but I have seen sites with much poorer design and content rank higher than this one. Any ideas welcome! Thanks in advance My site
Technical SEO | | mozUser14692366292850