Duplicate content and canonicalization confusion
-
Hello,
http://bit.ly/1b48Lmp and http://bit.ly/1BuJkUR pages have same content and their canonical refers to the page itself. Yet, they rank in search engines. Is it because they have been targeted to different geographical locations? If so, still the content is same.
Please help me clear this confusion.
Regards
-
I agree with you. It's all very confusing and little details make a BIG difference. Thanks for sticking with this.
-
Thanks a ton Donna for looking into the issue and helping at this level. I highly appreciate it
Their canonical tags confused me. As you have mentioned, the tags should have been one, I don't know why they are using two different ones. Probably, they have set the different geographic targets in Google Webmaster Tools and with the minor content variation and canonical tags, they want to signal Google to treat both the pages differently. I mean it's a big name in the world of ERP. They can't mess up with the canonical tags.
What do you think?
-
Okay. Let's start over looking at it from a goal perspective. I compared the two pages. Here is the difference between the two in terms of page text, highlighted in yellow - http://63.249.66.211/comparison.html. The differences are in the URL, the phone numbers at the top, a word here and there in the middle, and the 2nd block of text and photo under "Explore Our Solutions".
The first page, which I'll call India, has a canoncial tag pointing to itself. (http://www.sap.com/india/pc/bp/erp.html"/>) .
The second page, which I'll call UK, has a canoncial tag, also pointing to itself. (http://www.sap.com/uk/pc/bp/erp.html"/>).
- If you want both pages to rank and have authority, then you use the canonical tag. You need to use the same canonical tag on both pages. Right now they're different. That will essentially tell Google to treat the two pages as one; to show one or the other in search results, but considate their combined SEO value into one for ranking purposes.
- If you only want one page to rank, then noindex the other.
Does that make more sense?
-
Thanks for the reply Donna but my question is bit different. Could you please take a look at the rel canonical tag of the urls I posted. The content on both the pages is 100% same. The only difference is that they are targeted at different geographic locations. The canonical tags point to the page itself and not any master page.
-
This might help Shailendra - https://support.google.com/webmasters/answer/139066?hl=en. Skim down to (or search for) the part beginning with "This indicates the preferred URL", about half-way down the page.
Bottom line, Google attempts to respect canonical tags but it's no guarantee. Increase your chances by using "absolute paths rather than relative paths with the
rel="canonical"
link element". -
Thanks everyone for the response! But I am still confused. The two links that I have posted in my initial question have exactly the same content on both the pages (targeted at different geographic locations) and their canonical tags do not refer to any master page but to them itself, i.e. canonical tag on page A refers to A and canonical tag on page B refers to B. Please take a look at both the pages: http://bit.ly/1b48Lmp and http://bit.ly/1BuJkUR
Regards
-
Canonical pages still get indexed at Google's discretion.
A related question was asked in March 2013 that I think, explains what you're seeing. I've cut and pasted the relevant part below. Mememax is the author.
"Normally the only thing which will prevent a page from ranking is noindex tag. If you don't want to have it indexed just noindex it, if that page has been laready indexed, put the noindex tag and delete from index using GWT option.
Concerning the canonical tag thing, it will consolidate the seo value in one page but it won't prevent those page to appear in rankings, however you may have two cases:
-
the two or more pages are identical. In that case google may accept the canonicalization and show always the original page.
-
the two or more pages are slightly different, it's the case of paginated pages which are canonicalized using rel next/prev. In that sense the whole value will be consolidated in page 1 but then the page which will be shown in the rankings will be the one which responds to that query, for example if someone is looking for blue glass, google will return the page which shows blue glass listing if that's different from the first one."
-
-
Yes, if they were directly competing against each other, you'd expect one of them to drop out of the rankings. What are they both ranking for?
If they are both showing up in the same search, my guess would be that they are very new and Google hasn't noticed the duplication.
But if you see the ranking in different searches (like Google UK and Google India), then you are probably right, Google does not see them as duplicate since they are being shown to different audiences.
-
Hi,
I am sharing two Matt cutts video on this to clear your confusion.I hope it helps.
https://www.youtube.com/watch?v=GFf1gwr6HJw
Thanks
Got a burning SEO question?
Subscribe to Moz Pro to gain full access to Q&A, answer questions, and ask your own.
Browse Questions
Explore more categories
-
Moz Tools
Chat with the community about the Moz tools.
-
SEO Tactics
Discuss the SEO process with fellow marketers
-
Community
Discuss industry events, jobs, and news!
-
Digital Marketing
Chat about tactics outside of SEO
-
Research & Trends
Dive into research and trends in the search industry.
-
Support
Connect on product support and feature requests.
Related Questions
-
Duplicate Content and Subdirectories
Hi there and thank you in advance for your help! I'm seeking guidance on how to structure a resources directory (white papers, webinars, etc.) while avoiding duplicate content penalties. If you go to /resources on our site, there is filter function. If you filter for webinars, the URL becomes /resources/?type=webinar We didn't want that dynamic URL to be the primary URL for webinars, so we created a new page with the URL /resources/webinar that lists all of our webinars and includes a featured webinar up top. However, the same webinar titles now appear on the /resources page and the /resources/webinar page. Will that cause duplicate content issues? P.S. Not sure if it matters, but we also changed the URLs for the individual resource pages to include the resource type. For example, one of our webinar URLs is /resources/webinar/forecasting-your-revenue Thank you!
Technical SEO | | SAIM_Marketing0 -
404 Error Pages being picked up as duplicate content
Hi, I recently noticed an increase in duplicate content, but all of the pages are 404 error pages. For instance, Moz site crawl says this page: https://www.allconnect.com/sc-internet/internet.html has 43 duplicates and all the duplicates are also 404 pages (https://www.allconnect.com/Coxstatic.html for instance is a duplicate of this page). Looking for insight on how to fix this issue, do I add an rel=canonical tag to these 60 error pages that points to the original error page? Thanks!
Technical SEO | | kfallconnect0 -
Should you use the canonicalization tag when the content isn't exactly a duplicate?
We have a site that pull data from different sources with unique urls onto a main page and we are thinking about using the canonicalization tag to keep those source pages from being indexed and to give any authority to the main page. But this isn’t really what canonicalization is supposed to be used for so I’m unsure of if this is the right move.
Technical SEO | | Fuel
To give some more detail: We manage a site that has pages for individual golf courses. On the golf course page in addition to other general information we have sections on that page that show “related articles” and “course reviews”.
We may only show 4 or 5 on each of those courses pages per page, but we have hundreds of related articles and reviews for each course. So below “related articles” on the course page we have a link to “see more articles” that would take the user to a new page that is simply a aggregate page that houses all the article or review content related to that course.
Since we would rather have the overall course page rank in SERPs rather than the page that lists these articles, we are considering canonicalizing the aggregate news page up to the course page.
But, as I said earlier, this isn’t really what the canonicalization tag is intended for so I’m hesitant.
Has anyone else run across something like this before? What do you think?0 -
Duplicate Content Issues
We have some "?src=" tag in some URL's which are treated as duplicate content in the crawl diagnostics errors? For example, xyz.com?src=abc and xyz.com?src=def are considered to be duplicate content url's. My objective is to make my campaign free of these crawl errors. First of all i would like to know why these url's are considered to have duplicate content. And what's the best solution to get rid of this?
Technical SEO | | RodrigoVaca0 -
Duplicate page content - index.html
Roger is reporting duplicate page content for my domain name and www.mydomain name/index.html. Example: www.just-insulation.com
Technical SEO | | Collie
www.just-insulation.com/index.html What am I doing wrongly, please?0 -
URL query considered duplicate content?
I have a Magento site. In order to reduce duplicate content for products of the same style but with different colours I have combined them on to 1 product page. I would like to allow the pictures to be dynamic, i.e. allow a user to search for a colour and all the products that offer that colour appear in the results, but I dont want the default product image shown but the product image for that colour applying to the query. Therefore to do this I have to append a query string to the end of the URL to produce this result: www.website.com/category/product-name.html?=red My question is, will the query variations then be picked up as duplicate content: www.website.com/category/product-name.html www.website.com/category/product-name.html?=red www.website.com/category/product-name.html?=yellow Google suggest it has contingencies in its algorithm and I will not be penalised: http://googlewebmastercentral.blogspot.co.uk/2007/09/google-duplicate-content-caused-by-url.html But other sources suggest this is not accurate. Note the article was written in 2007.
Technical SEO | | BlazeSunglass0 -
Duplicate Home Page content and title ... Fix with a 301?
Hello everybody, I have the following erros after my first crawl: Duplicate Page Content http://www.peruviansoul.com http://www.peruviansoul.com/ http://www.peruviansoul.com/index.php?id=2 Duplicate Page title http://www.peruviansoul.com http://www.peruviansoul.com/ http://www.peruviansoul.com/index.php?id=2 Do you think I could fix them redirecting to http://www.peruviansoul.com with a couple of 301 in the .htaccess file? Thank you all for you help. Gustavo
Technical SEO | | peruviansoul0 -
How to publish duplicate content legitimately without Panda problems
Let's imagine that you own a successful website that publishes a lot of syndicated news articles and syndicated columnists. Your visitors love these articles and columns but the search engines see them as duplicate content. You worry about being viewed as a "content farm" because of this duplicate content and getting the Panda penalty. So, you decide to continue publishing the content and use... <meta name="robots" content="noindex, follow"> This allows you do display the content for your visitors but it should stop the search engines from indexing any pages with this code. It should also allow robots to spider the pages and pass link value through them. I have two questions..... If you use "noindex" will that be enough to prevent your site from being considered as a content farm? Is there a better way to continue publication of syndicated content but protect the site from duplicate content problems?
Technical SEO | | EGOL0