How rel=canonical works with index, noindex ?
-
Hello all,
I had always wondered how the index,noindex affects to the canonical. And also if the canonical post should be included in the sitemap or not.
I posted this
http://www.comparativadebancos.co...
and with a rel=canonical to this that was published at the beginning of the month
http://www.comparativadebancos.co...
but then I have the first one in google
http://www.google.com/search?aq=f...
May be this is evident for you but, what is really doing the canonical? If I publish something with the canonical pointing to another page, will it still be indexed by google but with no penalty for duplicate content? Or the usual behaviour should have been to havent indexed the first post but just the second one?
Should I also place a noindex in the first post in addition to the canonical?
What am I missing here?
thanks
-
Antonio,
I came into this question a little late so I'm not sure how it was back when you asked it, but right now the problem I see is that the page that does exist ( http://www.comparativadebancos.com/mejores-depositos-bancarios-de-marzo-de-2011/ ) has a rel canonical tag pointing to the page that doesn't exist ( http://www.comparativadebancos.com/depositos/marzo/ ), which returns a 404 response code.
I think right now the best thing you can do would be to change the rel canonical tag on /mejores-depositos-bancarios-de-marzo-de-2011/ to be http://www.comparativadebancos.com/mejores-depositos-bancarios-de-marzo-de-2011/ .
-
I im saying that it is important to Google to tell them more what you want to use as your content without possible parameter "/" "www" adding a duplicate content penalty to your website.
-
Hi,
I agree that it will not help you to too much with stolen content. Unless Google has indexed you 1st they would probably give you 1st rights to the disputed content. The reason I believe you are getting with such good results on Google a non-indexed URL or what should be nonindexed is Google indexes everything regardless and from what Matt Cutts said "According to Google, the canonical link element is not considered to be a directive, but a hint that the web crawler will "honor strongly" "
my belief is Google is throwing more honor to dealing with the canonical.
I hope I was of some help.
Sincerely,
Thomas Zickell
-
Blueprint, as far as I understand it can't really be used to prevent people stealing your content because you need to have to similar versions and place the tag pointing to the one that is of lesser value or that you don't want to come up in place of the original. Or are you saying if you find some of your content elsewhere offsite you can place a canonical link to it, and this will tell the spiders it is your content rather than theres?
Antonio, if you have placed the tag on the newer page pointing to the older page you are telling the spiders that the newer page is the preferred/more original content.
-
I would say that rel=canonical is one of the single most vital parts of a website no matter how it's Written or hosted all must be set up to appropriately take traffic and simply tell Google I'm not trying to duplicate my content here is my <link rel="canonical" href="http://www.example.com/" /> and that way if anyone does haven't come across your content and try to make it their own they will be the ones penalized for stealing it not you. Always put this tag in the page that you have created and the one that you want Google to understand is your copy of your website content here is some info from Matt Cutts at Google as well as Wikipedia hope I am of help
http://www.mattcutts.com/blog/rel-canonical-html-head/
A canonical link element is an HTML element that helps webmasters prevent duplicate content issues by specifying the "canonical", or "preferred", version of a web page<sup id="cite_ref-googleblog_0-0" class="reference">[1]</sup><sup id="cite_ref-1" class="reference">[2]</sup><sup id="cite_ref-2" class="reference">[3]</sup> as part of search engine optimization.
Duplicate content issues occur when the same content is accessible from multiple URLs.<sup id="cite_ref-3" class="reference">[4]</sup> For example, <tt>http://www.example.com/page.html</tt> would be considered by search engines to be an entirely different page to<tt>http://www.example.com/page.html?parameter=1</tt>, even though both URLs return the same content. Another example is essentially the same (tabular) content, but sorted differently.
In February 2009, Google, Yahoo and Microsoft announced support for the canonical link element, which can be inserted into the section of a web page, to allow webmasters to prevent these issues.<sup id="cite_ref-4" class="reference">[5]</sup> The canonical link element helps webmasters make clear to the search engines which page should be credited as the original.
According to Google, the canonical link element is not considered to be a directive, but a hint that the web crawler will "honor strongly".<sup id="cite_ref-googleblog_0-1" class="reference">[1]</sup>
While the canonical link element has its benefits, Matt Cutts, who is the head of Google's webspam team, has claimed that the search engine prefers the use of 301 redirects. Cutts claims the preference for redirects is because Google's spiders can choose to ignore a canonical link element if they feel it is more beneficial to do so.<sup id="cite_ref-5" class="reference">[6]</sup>
[edit]Examples of the
canonical
link element<link rel="canonical" href="http://www.example.com/" />
<link rel="canonical" href="http://www.example.com/page.html" />
<link rel="canonical" href="http://www.example.com/directory/page.html" /> ```
-
you should give it time to settle down in the SERPS ... the results are muddy for a while but your canonicals will eventually show up if they have been implemented correctly.
-
I have already done it but my question come after this one
Where Rand suggest me to do the canonical thing I am explaining here. So my doubt is why it is indexing the new post better than the old one and how this is supposed to work.
From my understanding and also from your link, if I use rel=canonical is the "canonical" url the one that has to be indexed and not the one with "rel=canonical" but it has not been my case and now I have both indexed...
Any suggestion?
-
Is it the opposite. The new one has a rel=canonical to the old one because it was written with the same content that the old one but then it appears in the index.
Then the new one has been indexed and I thought it wasnt going to be indexed. But at the same time it ranks much higger than the old one...
-
According to Google a rel=canonical is just a hint 9although they say they strongly honour it) - http://googlewebmastercentral.blogspot.com/2009/02/specify-your-canonical.html. This might explain why your old page is still showing up int he results.
Has your new page been indexed yet?
Got a burning SEO question?
Subscribe to Moz Pro to gain full access to Q&A, answer questions, and ask your own.
Browse Questions
Explore more categories
-
Moz Tools
Chat with the community about the Moz tools.
-
SEO Tactics
Discuss the SEO process with fellow marketers
-
Community
Discuss industry events, jobs, and news!
-
Digital Marketing
Chat about tactics outside of SEO
-
Research & Trends
Dive into research and trends in the search industry.
-
Support
Connect on product support and feature requests.
Related Questions
-
Will canonical solve this?
Hi all, I look after a website which sells a range of products. Each of these products has different applications, so each product has a different product page. For eg. Product one for x application Product one for y application Product one for z application Each variation page has its own URL as if it is a page of its own. The text on each of the pages is slightly different depending on the application, but generally very similar. If I were to have a generic page for product one, and add canonical tags to all the variation pages pointing to this generic page, would that solve the duplicate content issue? Thanks in advance, Ethan
Technical SEO | | Analoxltd0 -
Indexation and visibility problem
Hi I am working on a website (usarrestsearch org) for 6 months. I wrote about 100 pages full of good content. for some reason I see only 75% of the pages indexed in GWT. and Im having problems with SERP positions not rising. I suspect that it might be connected to the structure of the site. will appreciate any help thanks
Technical SEO | | holdportals0 -
Please let me know if I am in a right direction with fixing rel="canonical" issue?
While doing my website crawl, I keep getting the message that I have tons of duplicated pages.
Technical SEO | | kirupa
http://example.com/index.php and http://www.example.com/index.php are considered to be the duplicates. As I figured out this one: http://example.com/index.php is a canonical page, and I should point out this one: http://www.example.com/index.php to it. Could you please let me know if I will do a right thing if I put this piece of code into my index.php file?
? Or I should use this one:0 -
Site Not Being Indexed
Hey Everyone - I have a site that is being treated strangely by google (at least strange to me) The site has 24 pages in the sitemap - submitted to WMT'S over 30 days ago I've manually triggered google to crawl the homepage and all connecting links as well and submitted a couple individually. Google has been parked the indexing at 14 of the 24 pages. None of the unindexed URL's have Noindex or follow tags on them - they are clearly and easily linked to from other places on the site. The site is a brand new domain, has no manual penalty history and in my research has no reason to be considered spammy. 100% unique handwritten content I cannot figure out why google isn't indexing these pages. Has anyone encountered this before? Know any solutions? Thanks in advance.
Technical SEO | | CRO_first0 -
How do I eliminate indexed products?
Please help! We got clobbered by Penguin and are at risk of having to close down after 10 years. We have been trying to figure out why and believe now it might be because of duplicate content. We added 2" inserts in March (over 500): http://www.trophycentral.com/inserts1.html Even though each is a different products, SEOMOZ is saying they are considered duplicate content. Given the timing, we think this might be the cause, even though it is totally legitimate. Question - since these are now indexed and since we can't easily add content quickly, what is the best way to handle this situation? A no-index tag? Is there a way to let Google know that their algorithm is detroying legitimate businesses??
Technical SEO | | trophycentraltrophiesandawards0 -
Getting rid of duplicate content with rel=canonical
This may sound like a stupid question, however it's important that I get this 100% straight. A new client has nearly 6k duplicate page titles / descriptions. To cut a long story short, this is mostly the same page (or rather a set of pages), however every time Google visits these pages they get a different URL. Hence the astronomical number of duplicate page titles and descriptions. Now the easiest way to fix this looks like canonical linking. However, I want to be absolutely 100% sure that Google will then recognise that there is no duplicate content on the site. Ideally I'd like to 301 but the developers say this isn't possible, so I'm really hoping the canonical will do the job. Thanks.
Technical SEO | | RiceMedia0 -
REL Canonical Error
In my crawl diagnostics it showing a Rel=Canonical error on almost every page. I'm using wordpress. Is there a default wordpress problem that would cause this?
Technical SEO | | mmaes0 -
Can I noindex most of my site?
A large number of the pages on my site are pages that contain things like photos and maps that are useful to my visitors, but would make poor landing pages and have very little written content. My site is huge. Would it be benificial to noindex all of these?
Technical SEO | | mascotmike0